Glue Fewshot Eval

Evaluates few-shot text classification performance across multiple natural language understanding tasks. It probes a model's ability to generalize from extremely limited labeled examples (16 per class) by generating synthetic training data and fine-tuning a classifier. Use when the user wants to benchmark on GLUE, or asks about evaluating this task. Reports Average performance.

qhjqhj00 713c38e 3.5 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/glue-fewshot-eval commit 713c38e020

Frequently asked questions

npx skillmds add qhjqhj00/glue-fewshot-eval