Mmkc Bench Eval

Evaluates how large multimodal models (LMMs) handle factual knowledge conflicts between their internal parametric knowledge and external multimodal evidence. It probes both behavioral alignment (whether models follow internal knowledge or external context) and conflict detection capabilities across coarse- and fine-grained settings. Use when the user wants to benchmark on MMKC-Bench, or asks about evaluating this task. Reports Detection Accuracy.

qhjqhj00 974426f 4.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/mmkc-bench-eval commit 974426ffea

Frequently asked questions

npx skillmds add qhjqhj00/mmkc-bench-eval