Agent Skills: agent-evaluation

Design and implement comprehensive evaluation systems for AI agents. Use when building evals for coding agents, conversational agents, research agents, or computer-use agents. Covers grader types, benchmarks, 8-step roadmap, and production integration.

UncategorizedID: autohandai/community-skills/agent-evaluation

Install this agent skill to your local

pnpm dlx add-skill https://github.com/autohandai/community-skills/agent-evaluation

Skill Files

Browse the full folder contents for agent-evaluation.

Download Skill

Loading file tree…

Select a file to preview its contents.