Agent Skills: rag-eval

Iterate on RAG systems with structured evals instead of eyeballing. This skill should be used when the user is tuning a RAG pipeline — changing retrieval prompts, swapping models, adjusting chunking, or debugging poor answers — and wants a cheap, ranked set of experiments with cost tracking and structured feedback on the stack. Also use when the user asks "how do I know if my RAG is working?", "this RAG eval is burning money", or "what should I try next on retrieval?".

UncategorizedID: glebis/claude-skills/rag-eval

Install this agent skill to your local

pnpm dlx add-skill https://github.com/glebis/claude-skills/rag-eval

Skill Files

Browse the full folder contents for rag-eval.

Download Skill

Loading file tree…

Select a file to preview its contents.