RAG Eval

Evaluate your RAG pipeline quality using Ragas metrics (faithfulness, answer relevancy, context precision). PREREQUISITE: You must have a RAG system integrated with OpenClaw (e.g. vector DB + retrieval). Use when: (1) testing RAG answer quality after config changes, (2) checking for hallucinations in retrieved-context answers, (3) running batch regression tests on a golden dataset, (4) comparing RAG performance before/after embedding or chunking changes. NOT for: general LLM chat evaluation without retrieval context, code review, or non-RAG agent outputs.

johnalbertini14-glitch Updated 1 repo stars

File contents

johnalbertini14-glitch/openclaw-skills/tree/main/skills/jonathanjing/rag-eval commit a05292d4e6

Frequently asked questions

npx skillmds@latest add johnalbertini14-glitch/rag-eval