Regradient 160k Eval

Evaluates the ability of vision-language models to generate accurate and semantically rich chest X-ray radiology reports from medical images. It tests both lexical overlap and clinical semantic alignment of the generated findings and impressions against ground-truth reports. Use when the user wants to benchmark on ReXGradient-160K, or asks about evaluating this task. Reports COMET.

qhjqhj00 02ffccd 3.8 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/regradient-160k-eval commit 02ffccd605

Frequently asked questions

npx skillmds add qhjqhj00/regradient-160k-eval