Squad V1.1 Eval

Measures extractive question answering capability by requiring the model to identify a text span in a passage that answers a given question. It tests precise token-level span prediction and contextual understanding. Use when the user wants to benchmark on SQuAD v1.1, or asks about evaluating this task. Reports exact-match (EM).

qhjqhj00 4516974 2.5 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/squad-v1.1-eval commit 45169744a2

Frequently asked questions

npx skillmds add qhjqhj00/squad-v1-1-eval