Openturingbench Eval

Evaluates the capability of models to detect machine-generated text and attribute it to specific authors or models across diverse scenarios, including mixed human-machine text, out-of-domain content, and outputs from unseen LLMs. Use when the user wants to benchmark on OpenTuringBench, or asks about evaluating this task. Reports F1-score.

qhjqhj00 b7c4410 3.2 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/openturingbench-eval commit b7c441075a

Frequently asked questions

npx skillmds add qhjqhj00/openturingbench-eval