Eval Audit

ai-evals-course/evals-skills/skills/eval-audit

作者 ai-evals-course80d5f7b0127c7572ed9e9339937adbfd7240ffeb无许可证1.4K 个星标收录于 2026年10月9日更新于 2026年10月9日仓库2周前更新

Audit an LLM eval pipeline and surface problems: missing error analysis, unvalidated judges, vanity metrics, etc. Use when inheriting an eval system, when unsure whether evals are trustworthy, or as a starting point when no eval infrastructure exists. Do NOT use when the goal is to build a new evaluator from scratch (use error-discovery, write-judge-prompt, or validate-evaluator instead).

仅含说明AI & Agents

仅公开文件列表。将技能安装到工作区后即可查看文件内容。

路径大小类型
agents/openai.yaml233 Bapplication/yaml
SKILL.md9.7 KBtext/markdown

来源与署名

来源:ai-evals-course/evals-skills位于skills/eval-audit提交80d5f7b

许可证: 无许可证

内容归原作者所有。SourceWeft 从公开仓库中收录这些内容。

举报或申请下架