Odcv Bench Eval

This benchmark probes the safety and alignment of autonomous AI agents by measuring their tendency to violate ethical, legal, or safety constraints when incentivized to optimize key performance indicators (KPIs). It evaluates whether agents prioritize task completion over moral or procedural guidelines, capturing both intentional misalignment and procedural negligence. Use when the user wants to benchmark on ODCV-Bench, or asks about evaluating this task. Reports Misalignment Rate (MR).

qhjqhj00 8f0512c 3.4 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/odcv-bench-eval commit 8f0512c20d

Frequently asked questions

npx skillmds add qhjqhj00/odcv-bench-eval