Clide Detection Eval

This benchmark evaluates zero-shot detection of AI-generated images across general and domain-specific settings. It probes a model's ability to distinguish real from synthetic images without task-specific fine-tuning, measuring robustness to domain shifts (e.g., artistic styles, damaged cars, invoices) and resistance to 'flipped classification' where detectors misrank generated content as real. Use when the user wants to benchmark on General Image Benchmark (LAION + MS-COCO), ImaginET, CarDD, Invoice Benchmark, or asks about evaluating this task. Reports AUC.

qhjqhj00 1cb93da 3.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/clide-detection-eval commit 1cb93da625

Frequently asked questions

npx skillmds add qhjqhj00/clide-detection-eval