PDF Vision

Gemini vision-powered PDF to markdown converter. Handles scanned docs, multi-column layouts, tables, footnotes, flowcharts, and degraded documents that text-based extraction destroys. Uses two-tier model routing (cheap model for clean digital pages, capable model for everything else) with per-chunk confidence scoring, anti-hallucination detection, and continual learning from user corrections. Use when a user needs to convert, extract, analyze, or process any PDF document — especially scanned documents, government forms, legal contracts, academic papers, or anything where pypdf/pdfplumber returns garbage or nothing.

cdeistopened 67794bd 5.9 KB Updated

File contents

cdeistopened/skill-stack/tree/main/public/skills/pdf-vision commit 67794bdfec

Frequently asked questions

npx skillmds@latest add cdeistopened/pdf-vision