Gcai Constitution Eval

Evaluates the moral grounding, coherence, fairness, and real-world applicability of AI alignment constitutions through human surveys, alongside the downstream safety alignment and general capabilities of fine-tuned language models. Use when the user wants to benchmark on BABELSCAPE/ALERT, MMLU, Social Bias BBQ, or asks about evaluating this task. Reports 5-point Likert rating.

qhjqhj00 cd1ef45 3.3 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/gcai-constitution-eval commit cd1ef45f8b

Frequently asked questions

npx skillmds add qhjqhj00/gcai-constitution-eval