Mdia Eval

This benchmark evaluates a model's ability to generate coherent, contextually appropriate, and lexically diverse dialogue responses across 46 languages. It specifically probes cross-lingual transfer capabilities and measures the performance gap between high-resource and low-resource languages in open-domain conversation. Use when the user wants to benchmark on MDIA, or asks about evaluating this task. Reports sacreBLEU.

qhjqhj00 c46819f 3.8 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/mdia-eval commit c46819f6d5

Frequently asked questions

npx skillmds add qhjqhj00/mdia-eval