Deepeval

confident-ai/deepeval/docs/public/.well-known/agent-skills/deepeval

作者 confident-ai0fb05d028b5ded9e4c466140ab3f3af90b801944Apache-2.018K 個星標收錄於 2026年10月9日更新於 2026年10月9日儲存庫昨天更新

DeepEval evaluation workflow for AI agents and LLM applications. TRIGGER when the user wants to evaluate or improve an AI agent, tool-using workflow, multi-turn chatbot, RAG pipeline, or LLM app; add evals; generate datasets or goldens; use deepeval generate; use deepeval test run; send results to Confident AI; monitor production; run online evals; inspect traces; or iterate on prompts, tools, retrieval, or agent behavior from eval failures. AI agents are the primary use case. Covers Python SDK, pytest eval suites, CLI generation, traced evals, Confident AI reporting, and agent-driven improvement loops. DO NOT TRIGGER for unrelated generic pytest, non-AI test setup, or non-DeepEval observability work unless the user asks to compare or migrate to DeepEval; for instrumenting an app with DeepEval tracing, @observe, or framework integrations (use the `deepeval-tracing` skill); or for raw OpenTelemetry / OTLP export without the deepeval package (use the `deepeval-otel` skill).

包含腳本AI & Agents
  1. 0fb05d028b5ded9e4c466140ab3f3af90b801944目前提交 0fb05d0發布於 2026年10月9日

來源與署名

來源:confident-ai/deepeval位於docs/public/.well-known/agent-skills/deepeval提交0fb05d0

授權條款: Apache-2.0

內容歸原作者所有。SourceWeft 從公開儲存庫中收錄這些內容。

檢舉或申請下架