Crawl Web Text

Extract clean, source-recorded text from public webpages or saved HTML files. Use when Codex needs to crawl or fetch public web pages for textual content, convert HTML into readable text, preserve source metadata, or prepare webpage text for summarization, analysis, citation, or downstream processing while respecting robots.txt, access limits, paywalls, logins, CAPTCHA, and anti-bot restrictions.

CoUse-123 Updated

File contents

CoUse-123/codex-skills/tree/main/skills/crawl-web-text commit b1faf2222b

Frequently asked questions

npx skillmds@latest add couse-123/crawl-web-text