Kvasir Vqa Eval

Probes a model's ability to perform multimodal understanding and generation on gastrointestinal endoscopic images. It evaluates capabilities in descriptive captioning, answering clinical questions about visual findings, and synthesizing anatomically plausible medical images from text prompts. Use when the user wants to benchmark on Kvasir-VQA, or asks about evaluating this task. Reports BLEU.

qhjqhj00 50c8098 3.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/kvasir-vqa-eval commit 50c80988d6

Frequently asked questions

npx skillmds add qhjqhj00/kvasir-vqa-eval