Neuclirbench Eval

Evaluates the ranking effectiveness of retrieval and reranking models across monolingual, cross-language, and multilingual information retrieval tasks. It specifically probes how well systems handle language mismatches and multilingual document collections without relying on simple keyword matching. Use when the user wants to benchmark on NeuCLIRBench, or asks about evaluating this task. Reports nDCG@20.

qhjqhj00 48ddf88 3.1 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/neuclirbench-eval commit 48ddf88bbb

Frequently asked questions

npx skillmds add qhjqhj00/neuclirbench-eval