PDF Extract Confidence

Extract text from PDF documents (digital vector, scanned, or hybrid) with normalized per-word confidence scores (0.0 to 1.0) and output dual-purpose JSON containing complete document text and word-level audit metadata. Use this skill whenever the user asks to extract text from a PDF, inspect or verify word recognition confidence scores, extract PDFs into JSON, or check for low-confidence or potentially misrecognized words in PDF files.

terravic Updated

File contents

terravic/pdf-extract-confidence-skill/tree/main/skills/pdf-extract-confidence commit 6adb4591c3

Frequently asked questions

npx skillmds@latest add terravic/pdf-extract-confidence