Indicmmlu Pro Eval

Evaluates large language models on multi-task language understanding across nine major Indic languages. It probes capabilities in reading comprehension, reasoning, and knowledge retention by adapting the English MMLU-Pro benchmark through machine translation and rigorous quality assurance. Use when the user wants to benchmark on IndicMMLU-Pro, or asks about evaluating this task. Reports Accuracy.

qhjqhj00 038cc58 2.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/indicmmlu-pro-eval commit 038cc5897e

Frequently asked questions

npx skillmds add qhjqhj00/indicmmlu-pro-eval