PDF Extraction Eval

Evaluates open-source PDF information extraction tools across multiple content elements (metadata, references, tables, paragraphs, sections, etc.) on academic documents. It probes how well different tools handle layout-based segmentation, text extraction, and structural recognition in real-world academic PDFs. Use when the user wants to benchmark on DocBank, or asks about evaluating this task. Reports F1 score.

qhjqhj00 6a3d5a0 3.1 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/pdf-extraction-eval commit 6a3d5a0304

Frequently asked questions

npx skillmds add qhjqhj00/pdf-extraction-eval