Agent Skills: ai-evals

Designs trustworthy LLM, agent, responsible-AI, and multimodal evaluations. Use when measuring quality, fairness, privacy, grounding, safety, or judge reliability.

UncategorizedID: vasilyu1983/ai-agents-public/ai-evals

Install this agent skill to your local

pnpm dlx add-skill https://github.com/vasilyu1983/ai-agents-public/ai-evals

Skill Files

Browse the full folder contents for ai-evals.

Download Skill

Loading file tree…

Select a file to preview its contents.