Web Crawling

Crawl websites to extract text content from multiple pages and internal links. Use when the user wants to scrape a website, gather content across a section, archive docs text, understand a site more thoroughly by following internal links, or follow internal links beyond a single page. Prefer this general skill for bounded crawl workflows, route to Crawl4AI or Katana when the task is specifically AI-ready ingestion or site mapping, and escalate blocked pages to camofox before any residential-proxy discussion.

valtterimelkko Updated

File contents

valtterimelkko/agent-workflow-skills/tree/main/skills/web-crawling commit 46918bbd5e

Frequently asked questions

npx skillmds@latest add valtterimelkko/web-crawling