Natural Instructions Eval

Evaluates a model's ability to generalize to unseen NLP tasks by leveraging crowdsourced natural language instructions alongside training data. It measures how well instruction-based learning transfers across different task categories, datasets, and individual tasks compared to data-only training. Use when the user wants to benchmark on Natural Instructions, or asks about evaluating this task. Reports ROUGE-L.

qhjqhj00 bf45bd1 3.2 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/natural-instructions-eval commit bf45bd1e7c

Frequently asked questions

npx skillmds add qhjqhj00/natural-instructions-eval