Refute Your Own Answer Blind

When an answer arrives suspiciously clean — especially on a high-stakes claim, a safety judgment, or anything your own confidence is trying to launder into truth — strip authorship and confidence, pay an adversary to destroy it, and require the attacks to fail on execution rather than on counter-prose. The failure it prevents: self-preference and plausibility bias, where the model protects its own first answer because it sounds coherent and every review quietly reuses that same coherence. Triggers on "this should work", policy/safety claims, migration plans, root-cause explanations, and answers whose per-step confidence is uniformly high.

XyraSinclair Updated

File contents

XyraSinclair/ideonomy/tree/main/skills/refute-your-own-answer-blind commit 3b826f19db

Frequently asked questions

npx skillmds@latest add xyrasinclair/refute-your-own-answer-blind