Squad Fquad Eval

Evaluates a model's ability to extract and rank answer candidates from a given context for phrase-indexed question answering. It probes both the quality of candidate retrieval and the accuracy of final answer selection against gold spans. Use when the user wants to benchmark on SQuAD v1.1, FQuAD, or asks about evaluating this task. Reports exact-match.

qhjqhj00 a66f01b 2.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/squad-fquad-eval commit a66f01b793

Frequently asked questions

npx skillmds add qhjqhj00/squad-fquad-eval