
Evals Start
ai-evals-course/evals-skills/skills/evals-startby ai-evals-course80d5f7b0127c7572ed9e9339937adbfd7240ffebNo license1.4K starsListed Oct 9, 2026Updated Oct 9, 2026Repository updated 2 weeks ago
Entry point for evals. Use when the user asks for help with evals, does not know where to begin, or asks for something no other skill in this plugin matches. Do NOT use when a more specific skill in this plugin already matches; load that skill directly.
Add to a SourceWeft workspace
- Open the skill in your dashboard and add it to a workspace.
- Enable it for the chats that should use it.
This skill is instructions only: it ships no scripts to execute.
Add to SourceWeftYou will be asked to sign in first, then taken straight to this skill.
Ask your agent to install it
Paste this prompt into Claude Code, Codex, Cursor or another agent that can run commands — or into SourceWeft chat. The agent reads this skill's install guide, shows you its source, license and scripts, and installs it with the SourceWeft CLI once you agree.
Read https://sourceweft.com/skills/gh-ai-evals-course-evals-skills-evals-start-skills-evals-start-c95b34e01b160c57/install.md and install the skill it describes. Before installing, show me its source, license and whether it ships scripts, and wait for my OK. Ask me before changing anything else on my machine.Install it yourself from a terminal
For Claude Code, Codex, Cursor and other local agents. The SourceWeft CLI fetches the skill from its source repository at the commit scanned here, and verifies every file against the hashes recorded when the skill was scanned. If anything differs, nothing is written.
npx @sourceweft/cli skills install @ai-evals-course/evals-startAdd --agent claude-code, codex, cursor or universal to choose which agent gets it (Claude Code by default).
Upstream installer — not verified by SourceWeft
The open-source skills installer fetches the same pinned commit, but does not check the files against the hashes SourceWeft recorded.
npx skills add https://github.com/ai-evals-course/evals-skills/tree/80d5f7b0127c7572ed9e9339937adbfd7240ffeb/skills/evals-startSource and attribution
Source:ai-evals-course/evals-skillsinskills/evals-startat commit80d5f7b
License: No license
Content belongs to its original authors. SourceWeft indexes it from a public repository.
More from ai-evals-course/evals-skills

Validate Evaluator
ai-evals-course
Calibrates an LLM judge against human labels using data splits, TPR/TNR metrics, and bias correction.

Generate Synthetic Data
ai-evals-course
Generates diverse synthetic test inputs for LLM pipeline evaluation using dimension-based tuple generation.

Evaluate Rag
ai-evals-course
Guides evaluation of RAG retrieval and generation quality with metrics, datasets and chunking optimization.

Eval Audit
ai-evals-course
Audits an LLM eval pipeline and produces a prioritized findings report with concrete fixes.
More in AI & Agents

Discernment Nudge
anthropics
Appends 2-3 specific follow-up questions to substantive answers so users can check facts, reasoning and missing context.

Ppt Template Creator
anthropics
Turns a user's PowerPoint template into a reusable skill that generates branded presentations.

Skill Development
anthropics
Guides creation of Claude Code plugin skills, covering structure, descriptions, progressive disclosure and validation.

Plugin Structure
anthropics
Guides the structure, manifest, and component layout of Claude Code plugins.

Command Development
anthropics
Guides creation of Claude Code slash commands, covering structure, YAML frontmatter, arguments and plugin features.

Claude Md Improver
anthropics
Audits CLAUDE.md files in a repository, scores their quality, and applies approved targeted improvements.