muiE-eval

This benchmark evaluates a model's ability to perform universal information extraction (NER, RE, EE) and fine-grained cross-modal grounding (segmentation/tracking) across text, image, audio, and video modalities in a unified zero-shot setting. It probes the model's capacity to align semantic information with visual/auditory content and handle modality-shared versus modality-specific scenarios without task-specific fine-tuning. Use when the user wants to benchmark on MUIE, or asks about evaluating this task. Reports F1 (NER).

qhjqhj00 a25a5ea 4.4 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/muiE-eval commit a25a5ea80b

Frequently asked questions

npx skillmds add qhjqhj00/muie-eval