Dl21 Dl22 Ir Eval

This protocol evaluates information retrieval systems by measuring their ranking effectiveness on passage retrieval tasks using both original seed queries and LLM-generated query variants aligned with specific demographic or textual profiles. It probes whether retrieval systems perform consistently across diverse user personas and query transformations, revealing potential disparities in system behavior and ranking stability. Use when the user wants to benchmark on DL21 & DL22 (TREC Deep Learning Track), or asks about evaluating this task. Reports NDCG@10.

qhjqhj00 887d3a4 4.3 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/dl21-dl22-ir-eval commit 887d3a4f25

Frequently asked questions

npx skillmds add qhjqhj00/dl21-dl22-ir-eval