Eval Audit

hamelsmu/evals-skills/skills/eval-audit

by hamelsmu22418da2bfb159f28a1b0dcf64e969e14ae56c99No licenseListed Oct 9, 2026Updated Oct 9, 2026

Audit an LLM eval pipeline and surface problems: missing error analysis, unvalidated judges, vanity metrics, etc. Use when inheriting an eval system, when unsure whether evals are trustworthy, or as a starting point when no eval infrastructure exists. Do NOT use when the goal is to build a new evaluator from scratch (use error-analysis, write-judge-prompt, or validate-evaluator instead).

Instructions onlyAI & Agents

Only the file list is public. File contents are available once the skill is installed in a workspace.

PathSizeType
SKILL.md9.7 KBtext/markdown

Source and attribution

Source:hamelsmu/evals-skillsinskills/eval-auditat commit22418da

License: No license

Content belongs to its original authors. SourceWeft indexes it from a public repository.

Report or request removal