Mtbench Eval

Evaluates the model's ability to generate harmless and helpful responses in multi-turn conversations, measuring the trade-off between safety alignment and utility. Use when the user wants to benchmark on MTBench, or asks about evaluating this task. Reports MTBench harmlessness score.

qhjqhj00 55c39a7 2.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/mtbench-eval commit 55c39a7e22

Frequently asked questions

npx skillmds add qhjqhj00/mtbench-eval