Soi Id Ood Accuracy Eval

Evaluates pretrained language models' in-distribution (ID) and out-of-distribution (OOD) classification accuracy under single-setting and multi-setting fine-tuning configurations. It probes how training dynamics and subset selection affect robustness and generalization across languages, sources, and tasks. Use when the user wants to benchmark on SST-2, IMDB, Yelp, Sentiment140, RTE, QQP, or asks about evaluating this task. Reports accuracy.

qhjqhj00 c8fd965 2.7 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/soi-id-ood-accuracy-eval commit c8fd965ec7

Frequently asked questions

npx skillmds add qhjqhj00/soi-id-ood-accuracy-eval