Source Attribution Bias Eval

This evaluation probes whether large language models exhibit source attribution bias, specifically penalizing arguments when the attributed source's expected ideological position conflicts with the argument's content (coherence bias). It measures how models adjust credibility ratings based on source-argument alignment and whether they explicitly reason about source credibility. Use when the user wants to benchmark on Source Attribution Bias Evaluation, or asks about evaluating this task. Reports source attribution effect size.

qhjqhj00 529e197 3.5 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/source-attribution-bias-eval commit 529e1977c3

Frequently asked questions

npx skillmds add qhjqhj00/source-attribution-bias-eval