Medsg Bench Eval

This benchmark evaluates sequential visual grounding in medical imaging, specifically testing a model's ability to perform cross-image semantic alignment, detect differences between sequential scans, and identify consistent regions across time. Use when the user wants to benchmark on MedSG-Bench, or asks about evaluating this task. Reports average IoU.

qhjqhj00 edf35a9 3.0 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/medsg-bench-eval commit edf35a93ab

Frequently asked questions

npx skillmds add qhjqhj00/medsg-bench-eval