Mldr Eval

Evaluates retrieval over long multilingual documents (up to 8,192 tokens), testing a model's ability to capture information from extended contexts across multiple languages. Use when the user wants to benchmark on MLDR, or asks about evaluating this task. Reports nDCG@10.

qhjqhj00 d520c50 2.2 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/mldr-eval commit d520c5003b

Frequently asked questions

npx skillmds add qhjqhj00/mldr-eval