Visual Information Extraction Eval

Evaluates a model's ability to extract entity spans and link them to key-value pairs from complex, real-world document images. It probes joint vision-language understanding, handling poor image quality, occlusion, and multi-lingual text without relying on external OCR pipelines. Use when the user wants to benchmark on FUNSD, XFUND, CORD, SIBR, or asks about evaluating this task. Reports F1-score.

qhjqhj00 0be3a4a 3.3 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/visual-information-extraction-eval commit 0be3a4ae54

Frequently asked questions

npx skillmds add qhjqhj00/visual-information-extraction-eval