Reranker Benchmark Eval

Evaluates the ability of embedding and reranking models to retrieve and rank relevant paragraphs from descriptive linguistic grammars based on typological feature queries, specifically testing their capacity to filter noisy or partially relevant context. Use when the user wants to benchmark on The Benchmark for Rerankers, or asks about evaluating this task. Reports NDCG@k.

qhjqhj00 5caaa70 3.3 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/reranker-benchmark-eval commit 5caaa70dfc

Frequently asked questions

npx skillmds add qhjqhj00/reranker-benchmark-eval