Trafilatura

Extract clean article text and metadata from URLs or HTML with trafilatura CLI. Use for single-page extraction, piped/local HTML, bounded discovery. NOT for research synthesis (research), PDFs (docling), raw fetch (fetch), video (yt-dlp).

wyattowalsh Updated

File contents

wyattowalsh/agents/tree/main/skills/trafilatura commit 9927d39a95

Frequently asked questions

npx skillmds@latest add wyattowalsh/trafilatura