Eval Audit

ai-evals-course/evals-skills/skills/eval-audit

作者 ai-evals-course80d5f7b0127c7572ed9e9339937adbfd7240ffeb無授權條款1.4K 個星標收錄於 2026年10月9日更新於 2026年10月9日儲存庫2 週前更新

Audit an LLM eval pipeline and surface problems: missing error analysis, unvalidated judges, vanity metrics, etc. Use when inheriting an eval system, when unsure whether evals are trustworthy, or as a starting point when no eval infrastructure exists. Do NOT use when the goal is to build a new evaluator from scratch (use error-discovery, write-judge-prompt, or validate-evaluator instead).

僅含說明AI & Agents

僅公開檔案列表。將技能安裝到工作區後即可檢視檔案內容。

路徑大小類型
agents/openai.yaml233 Bapplication/yaml
SKILL.md9.7 KBtext/markdown

來源與署名

來源:ai-evals-course/evals-skills位於skills/eval-audit提交80d5f7b

授權條款: 無授權條款

內容歸原作者所有。SourceWeft 從公開儲存庫中收錄這些內容。

檢舉或申請下架