
Deepeval
作者 confident-aic144abbce848Apache-2.018K 個星標收錄於 2026年10月8日更新於 2026年10月8日儲存庫昨天更新
DeepEval evaluation workflow for AI agents and LLM applications. TRIGGER when the user wants to evaluate or improve an AI agent, tool-using workflow, multi-turn chatbot, RAG pipeline, or LLM app; add evals; generate datasets or goldens; use deepeval generate; use deepeval test run; send results to Confident AI; monitor production; run online evals; inspect traces; or iterate on prompts, tools, retrieval, or agent behavior from eval failures. AI agents are the primary use case. Covers Python SDK, pytest eval suites, CLI generation, traced evals, Confident AI reporting, and agent-driven improvement loops. DO NOT TRIGGER for unrelated generic pytest, non-AI test setup, or non-DeepEval observability work unless the user asks to compare or migrate to DeepEval; for instrumenting an app with DeepEval tracing, @observe, or framework integrations (use the `deepeval-tracing` skill); or for raw OpenTelemetry / OTLP export without the deepeval package (use the `deepeval-otel` skill).
- c144abbce848目前提交 c144abb發布於 2026年10月8日
來源與署名
來源:confident-ai/deepeval位於skills/deepeval提交c144abb
授權條款: Apache-2.0
內容歸原作者所有。SourceWeft 從公開儲存庫中收錄這些內容。
更多來自 confident-ai/deepeval 的技能
更多AI & Agents技能

Skill Development
anthropics
指導建立 Claude Code 外掛技能,涵蓋結構、描述、漸進式揭露與驗證。

Plugin Structure
anthropics
說明 Claude Code 外掛的結構、資訊清單與元件配置方式。

Command Development
anthropics
指導建立 Claude Code 斜線命令,涵蓋結構、YAML frontmatter、參數與外掛功能。

Claude Md Improver
anthropics
稽核儲存庫中的 CLAUDE.md 檔案、評估其品質,並在取得核准後套用針對性的改進。

Find Skills
vercel-labs
指導透過 npx skills 命令列工具探索、評估並安裝代理技能。

Plugin Creator
openai
為 Codex 建立外掛目錄,產生 .codex-plugin/plugin.json 清單與可選的外掛市集項目。