RAG Eval

Iterate on RAG systems with structured evals instead of eyeballing. This skill should be used when the user is tuning a RAG pipeline — changing retrieval prompts, swapping models, adjusting chunking, or debugging poor answers — and wants a cheap, ranked set of experiments with cost tracking and structured feedback on the stack. Also use when the user asks "how do I know if my RAG is working?", "this RAG eval is burning money", or "what should I try next on retrieval?".

glebis d05373f 2 files · 11.9 KB Updated

File contents

glebis/claude-skills/tree/main/rag-eval commit d05373f8cb

Frequently asked questions

npx skillmds@latest add glebis/rag-eval