Context Conflict Merge Eval

Evaluates how language models merge conflicting generated and retrieved contexts in open-domain QA. It probes whether models exhibit a systematic bias toward generated contexts over retrieved ones when only one context contains the correct answer. Use when the user wants to benchmark on NQ-CC, TQA-CC, or asks about evaluating this task. Reports DiffGR.

qhjqhj00 97c5867 3.1 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/context-conflict-merge-eval commit 97c58673d3

Frequently asked questions

npx skillmds add qhjqhj00/context-conflict-merge-eval