Thai PDF Extract

แปลง PDF ภาษาไทย (โดยเฉพาะเอกสารราชการ/เล่มมาตรฐาน) ให้เป็นไฟล์ text สะอาด โดยสระและวรรณยุกต์ไม่เพี้ยน ไม่แตกบรรทัด — ใช้ pdftotext -layout (poppler) แทนตัวแปลงทั่วไป (markitdown / PyPDF / pdfplumber) ที่ทำภาษาไทยพัง Use this skill whenever the user asks to read, extract, convert, ingest, or summarize a PDF that contains Thai text — including Thai government documents (มสพร./มรด./ประกาศ), standards drafts, meeting documents, or any PDF being ingested into a knowledge base, wiki, Obsidian vault, or corpus for later search. Also use it when a previous PDF extraction produced garbled Thai (สระลอย, วรรณยุกต์หลุด, คำอย่าง "กำร" แทน "การ") — this skill detects and repairs that font-encoding corruption. Do NOT use for PDFs with no Thai text (a plain English PDF needs no special handling).

DGA-SD Updated

File contents

DGA-SD/Claude-SKILL/tree/main/thai-pdf-extract commit 2d8b164498

Frequently asked questions

npx skillmds@latest add dga-sd/thai-pdf-extract