Medq Deg Eval

Evaluates multimodal large language models' robustness and metacognitive reliability when processing medical images with various quality degradations (e.g., blur, noise, motion, artifacts) across different clinical capability dimensions. Use when the user wants to benchmark on MedQ-Deg, or asks about evaluating this task. Reports accuracy.

qhjqhj00 789a44c 2.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/medq-deg-eval commit 789a44c201

Frequently asked questions

npx skillmds add qhjqhj00/medq-deg-eval