Hallucination Evaluator

Detect and measure ungroundedness in LLM and RAG outputs — claims the source doesn't support — by decomposing answers into atomic claims and checking each for entailment, so you can quantify faithfulness and gate on it instead of eyeballing it. Use when a RAG/LLM feature makes confident wrong claims, before shipping anything that must be factual, or to add a groundedness gate to evals/CI.

imtiazrayhan Updated

File contents

imtiazrayhan/agentscamp-library/tree/main/skills/hallucination-evaluator commit d07cb616c2

Frequently asked questions

npx skillmds@latest add imtiazrayhan/hallucination-evaluator