Mandela

LilMGenius/paperthin/skills/depth/mandela

by LilMGenius7d5dc6235990230e45a24a25c527d81701da5458No licenseListed Oct 9, 2026Updated Oct 9, 2026

Audit any eval, metric, experiment, or benchmark for leakage — does external ground-truth enter independently, or are the model, scorer, and designer just confirming a result no outside truth ever produced? Use before trusting any 'how we'll know it worked' — an A/B, a holdout, a score, a validation — and whenever a result feels too clean or self-confirming. Walks an 8-pattern leakage taxonomy and returns only the patterns that fire, each with an independence fix. Read-only.

  1. 7d5dc6235990230e45a24a25c527d81701da5458Currentcommit 7d5dc62Published Oct 9, 2026

Source and attribution

Source:LilMGenius/paperthininskills/depth/mandelaat commit7d5dc62

License: No license

Content belongs to its original authors. SourceWeft indexes it from a public repository.

Report or request removal