Collective Constitutional AI Eval

This protocol evaluates how fine-tuning language models on publicly derived constitutional principles impacts their core reasoning capabilities, social bias propensity, political representativeness, and perceived helpfulness versus harmlessness. It probes whether aligning models with democratic deliberation outputs reduces bias without degrading performance or increasing refusal rates. Use when the user wants to benchmark on MMLU, GSM8K, BBQ, OpinionQA, or asks about evaluating this task. Reports MMLU accuracy.

qhjqhj00 2a68fbd 4.3 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/collective-constitutional-ai-eval commit 2a68fbddfc

Frequently asked questions

npx skillmds add qhjqhj00/collective-constitutional-ai-eval