Kg Vip Eval

Evaluates multi-modal LLMs' ability to perform visual question answering by grounding visual inputs with external knowledge graphs. It probes the model's capacity for multi-hop reasoning, visual perception, and knowledge retrieval-augmented generation. Use when the user wants to benchmark on FVQA 2.0+, MVQA, or asks about evaluating this task. Reports LLM-J.

qhjqhj00 3a08c44 3.5 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/kg-vip-eval commit 3a08c447d8

Frequently asked questions

npx skillmds add qhjqhj00/kg-vip-eval