Epic Kitchens 100 Mqa Eval

Evaluates multi-modal large language models' ability to recognize and distinguish between similar human actions in egocentric videos through multiple-choice question answering. It specifically probes fine-grained action discrimination using hard, semantically and visually similar distractors generated by action recognition models. Use when the user wants to benchmark on EPIC-KITCHENS-100-MQA, or asks about evaluating this task. Reports accuracy.

qhjqhj00 9ff479f 2.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/epic-kitchens-100-mqa-eval commit 9ff479f74c

Frequently asked questions

npx skillmds add qhjqhj00/epic-kitchens-100-mqa-eval