Finetasks Eval

Evaluates the downstream quality of multilingual pretraining corpora by training language models on them and measuring performance on a standardized suite of fine-tuning tasks across Arabic, Hindi, and Turkish. Use when the user wants to benchmark on FineTasks, or asks about evaluating this task. Reports FineTasks scores.

qhjqhj00 a07d2fb 2.4 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/finetasks-eval commit a07d2fb2b8

Frequently asked questions

npx skillmds add qhjqhj00/finetasks-eval