Convert PDFs and document images into agent-ready Markdown with Docext

Run Docext locally or with a chosen model backend to turn PDFs and document images into structured Markdown for RAG, extraction, and review workflows.

agentskillexchange Updated 28 repo stars

File contents

Convert PDFs and document images into agent-ready Markdown with Docext

Run Docext locally or with a chosen model backend to turn PDFs and document images into structured Markdown for RAG, extraction, and review workflows.

Prerequisites

Python 3.11 environment, Docext package, source PDFs or document images, and a supported VLM backend such as vLLM, Ollama, or a configured hosted model provider

Installation

Basic usage or getting-started notes:

Documentation

Source

agentskillexchange/skills/tree/main/skills/convert-pdfs-and-document-images-into-agent-ready-markdown-with-docext commit 079ab370d7

Frequently asked questions

npx skillmds@latest add agentskillexchange/convert-pdfs-and-document-images-into-agent-ready-markdown-w