Fastfact Eval

Evaluates the long-form factuality of LLM-generated responses by extracting claims, verifying them against evidence, and comparing the system's factuality scores against human annotations. Use when the user wants to benchmark on FaStfact-Bench, or asks about evaluating this task. Reports F₁@K′.

qhjqhj00 962fedd 2.7 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/fastfact-eval commit 962feddbe1

Frequently asked questions

npx skillmds add qhjqhj00/fastfact-eval