Data Similarity Performance Eval

Evaluates whether distributional or embedding similarity between a model's pretraining data and downstream tasks predicts few-shot or finetuned performance. It probes the 'similarity hypothesis' by measuring correlations between aggregate and example-level text similarities and model accuracy. Use when the user wants to benchmark on BIG-bench Lite, GLUE, or asks about evaluating this task. Reports correlation coefficient.

qhjqhj00 549893f 3.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/data-similarity-performance-eval commit 549893f3d1

Frequently asked questions

npx skillmds add qhjqhj00/data-similarity-performance-eval