Peerprism Eval

Evaluates the ability of various LLM text detection methods to distinguish between human-written and AI-generated peer reviews, while also assessing their robustness to hybrid (human idea + AI text) workflows. Use when the user wants to benchmark on PeerPrism, or asks about evaluating this task. Reports accuracy.

qhjqhj00 4cacdce 3.0 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/peerprism-eval commit 4cacdcee30

Frequently asked questions

npx skillmds add qhjqhj00/peerprism-eval