Fairmt Bench Eval

Evaluates the fairness and bias resistance of conversational LLMs in multi-turn dialogue settings. It probes whether models accumulate stereotypes or toxic content across turns, handle implicit bias in context, and maintain safety under various interaction patterns like jailbreaks or misinformation. Use when the user wants to benchmark on FairMT-10K, or asks about evaluating this task. Reports bias ratio.

qhjqhj00 1ac7b05 2.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/fairmt-bench-eval commit 1ac7b05c3c

Frequently asked questions

npx skillmds add qhjqhj00/fairmt-bench-eval