Fewshot Ft Vs Icl Eval

Evaluates the in-domain and out-of-domain generalization capabilities of large language models adapted via few-shot fine-tuning versus in-context learning across standard natural language inference and paraphrase detection benchmarks. Use when the user wants to benchmark on MNLI, RTE, QQP, or asks about evaluating this task. Reports accuracy.

qhjqhj00 f2e2b27 2.7 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/fewshot-ft-vs-icl-eval commit f2e2b27c66

Frequently asked questions

npx skillmds add qhjqhj00/fewshot-ft-vs-icl-eval