Constitution AI Feedback Eval

This evaluation probes how different instructional guidelines (constitutions) shape AI-generated medical dialogues across specific socio-communicative dimensions like empathy, information gathering, and decision-making. It measures human preference for dialogue quality under varying constitutional constraints. Use when the user wants to benchmark on Custom AI-generated medical dialogues, or asks about evaluating this task. Reports Bradley-Terry preference rate.

qhjqhj00 7e11b45 3.0 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/constitution-ai-feedback-eval commit 7e11b4526f

Frequently asked questions

npx skillmds add qhjqhj00/constitution-ai-feedback-eval