Eval Audit

ai-evals-course/evals-skills/skills/eval-audit

by ai-evals-course80d5f7b0127c7572ed9e9339937adbfd7240ffebNo license1.4K starsListed Oct 9, 2026Updated Oct 9, 2026Repository updated 2 weeks ago

Audit an LLM eval pipeline and surface problems: missing error analysis, unvalidated judges, vanity metrics, etc. Use when inheriting an eval system, when unsure whether evals are trustworthy, or as a starting point when no eval infrastructure exists. Do NOT use when the goal is to build a new evaluator from scratch (use error-discovery, write-judge-prompt, or validate-evaluator instead).

Instructions onlyAI & Agents

Only the file list is public. File contents are available once the skill is installed in a workspace.

PathSizeType
agents/openai.yaml233 Bapplication/yaml
SKILL.md9.7 KBtext/markdown

Source and attribution

Source:ai-evals-course/evals-skillsinskills/eval-auditat commit80d5f7b

License: No license

Content belongs to its original authors. SourceWeft indexes it from a public repository.

Report or request removal