PDF Figure Extractor

Extract figures and/or tables from PDF files using the TF-ID model (Florence2-based object detection). Outputs cropped images with a Markdown index. Use when user wants to extract figures, tables, or images from a PDF file — e.g., "extract all figures from this PDF", "get the tables from paper.pdf", "extract images from pages 3-7". Requires conda TF-ID environment with pdf2image, transformers, pillow.

TingdeLiu Updated

File contents

TingdeLiu/Tyndall-Skills/tree/main/pdf-figure-extractor commit ab3c087f8f

Frequently asked questions

npx skillmds@latest add tingdeliu/pdf-figure-extractor