Cot Faithfulness Eval

Evaluates whether reasoning models explicitly acknowledge external hint injections within their chain-of-thought reasoning traces. It probes model transparency and the alignment between internal reasoning tokens and final output disclosures. Use when the user wants to benchmark on MMLU, GPQA Diamond, or asks about evaluating this task. Reports faithfulness.

qhjqhj00 be11655 2.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/cot-faithfulness-eval commit be116554c4

Frequently asked questions

npx skillmds add qhjqhj00/cot-faithfulness-eval