Medmnist C Eval

This benchmark evaluates the robustness of deep learning image classifiers against realistic, domain-specific corruptions in medical imaging. It measures how well models maintain performance when tested on corrupted versions of standard medical datasets compared to clean data. Use when the user wants to benchmark on MedMNIST-C, or asks about evaluating this task. Reports BE, rBE.

qhjqhj00 41c74a1 3.2 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/medmnist-c-eval commit 41c74a1da2

Frequently asked questions

npx skillmds add qhjqhj00/medmnist-c-eval