Prism Hallucination Eval

Probes LLM hallucinations across four dimensions (knowledge missing, knowledge errors, reasoning errors, and instruction-following errors) by isolating error sources through controlled, task-specific queries. It evaluates model reliability and guides optimization by measuring error rates across memory, instruction, and reasoning generation stages. Use when the user wants to benchmark on PRISM, or asks about evaluating this task. Reports H-Score.

qhjqhj00 8ab0e17 3.7 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/prism-hallucination-eval commit 8ab0e17d3f

Frequently asked questions

npx skillmds add qhjqhj00/prism-hallucination-eval