Crosslingual Mtf Eval

Evaluates zero-shot crosslingual generalization of multilingual LLMs after multitask finetuning. Probes language-agnostic task understanding, robustness to prompt translation, and scaling behavior across NLU, generative, and code tasks. Use when the user wants to benchmark on XNLI, XCOPA, XStoryCloze, XWinograd, HumanEval, or asks about evaluating this task. Reports accuracy.

qhjqhj00 0cb5905 3.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/crosslingual-mtf-eval commit 0cb5905c59

Frequently asked questions

npx skillmds add qhjqhj00/crosslingual-mtf-eval