Agent Skills: agent-evals

Design and implement evaluation frameworks for AI agents. Use when testing agent reasoning quality, building graders, doing error analysis, or establishing regression protection. Framework-agnostic concepts that apply to any SDK.

UncategorizedID: majiayu000/claude-skill-registry/agent-evals

Install this agent skill to your local

pnpm dlx add-skill https://github.com/majiayu000/claude-skill-registry/agent-evals

Skill Files

Browse the full folder contents for agent-evals.

Download Skill

Loading file tree…

Select a file to preview its contents.