PDF Extract

Extract text from PDF files for LLM processing

modbender Updated 12 repo stars

File contents

PDF Extract

Extract text from PDF files for LLM processing. Uses pdftotext from the poppler-utils package to convert PDF documents into plain text.

Commands

# Extract all text from a PDF
pdf-extract "document.pdf"

# Extract text from specific pages
pdf-extract "document.pdf" --pages 1-5

Install

sudo dnf install poppler-utils

modbender/skill-library-mcp/tree/main/data/pdf-extract commit 56e20b9153

Frequently asked questions

npx skillmds@latest add modbender/pdf-extract