Dmid Mammography Report Eval

Evaluates a vision-language model's ability to generate clinically accurate and linguistically fluent mammography reports from multi-view breast images. It probes both natural language generation quality and domain-specific diagnostic reasoning, specifically BI-RADS categorization and breast density assessment. Use when the user wants to benchmark on DMID, or asks about evaluating this task. Reports BI-RADS Accuracy.

qhjqhj00 ea0360a 3.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/dmid-mammography-report-eval commit ea0360a564

Frequently asked questions

npx skillmds add qhjqhj00/dmid-mammography-report-eval