Squad Eval

Measures a model's ability to extract precise answer spans from a given context paragraph in response to a natural language question, testing reading comprehension and span prediction. Use when the user wants to benchmark on SQuAD 1.1/2.0, or asks about evaluating this task. Reports F1.

qhjqhj00 f1501ef 2.4 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/squad-eval commit f1501ef848

Frequently asked questions

npx skillmds add qhjqhj00/squad-eval