Multilingual LLM Downstream Eval

Evaluates the downstream capabilities of multilingual LLMs trained on filtered pretraining data. It probes reading comprehension, general knowledge, natural language understanding, common-sense reasoning, and generative tasks across multiple languages. Use when the user wants to benchmark on FineTasks, SmolLM tasks suite, or asks about evaluating this task. Reports average rank.

qhjqhj00 c4c5cee 3.0 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/multilingual-llm-downstream-eval commit c4c5cee4d5

Frequently asked questions

npx skillmds add qhjqhj00/multilingual-llm-downstream-eval