Bias Resistant Human Aligned Judging

Use this skill when a user wants evaluator data that probes whether a judge is fair, stable, and aligned with human preferences instead of being swayed by position, length, self-style, names, or misleading cues. Trigger it when people say things like 'test evaluator bias', 'make judge robustness data', 'see if the evaluator changes when I swap the order', or 'build hard examples that look different on the surface but should get the same verdict'. Plain-language examples include: 'check if the judge is biased toward longer answers', 'test self-preference or name bias', 'make order-swap stability data', and 'create fairness-style evaluator tests'.

dingxingdi Updated

File contents

dingxingdi/paper_fast_search_backup/tree/main/examples/evol_ability/20260325_170549/profiles/eval/skills/bias-resistant-human-aligned-judging commit 1581b70e6e

Frequently asked questions

npx skillmds@latest add dingxingdi/bias-resistant-human-aligned-judging-2