Hallucination Mitigation Eval

Evaluates the ability of Large Vision-Language Models to generate factually aligned outputs by measuring object hallucination rates in captions and yes/no answers, as well as logical reasoning and attribute consistency across diverse visual prompts. Use when the user wants to benchmark on POPE, CHAIR, MMHal-Bench, or asks about evaluating this task. Reports POPE Average Accuracy.

qhjqhj00 87fba75 3.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/hallucination-mitigation-eval commit 87fba752cc

Frequently asked questions

npx skillmds add qhjqhj00/hallucination-mitigation-eval