AnyCrawl Instagram Scraper
Use this skill when the user gives a public Instagram URL and wants the richest extraction available through AnyCrawl.
Workflow
- Accept a single public Instagram URL.
- Validate that it is one of the supported shapes:
https://www.instagram.com/<username>/https://www.instagram.com/reel/<shortcode>/https://www.instagram.com/p/<shortcode>/
- Run the bundled extractor:
node .claude/skills/anycrawl-instagram-scraper/scripts/run.js "<instagram-url>"
- Read the JSON file path printed by the script.
- Summarize the important fields for the user:
- page type
- owner/profile info
- caption text
- media URLs and media type
- engagement stats
- hashtags and mentions
- screenshot URL
- anything clearly blocked by login or consent walls
Output Handling
The script writes a full JSON artifact to /tmp by default and prints a compact summary to stdout.
The artifact includes:
normalized: structured Instagram-focused extractionanycrawl: raw AnyCrawl response, includingmarkdown,html,links, andscreenshot@fullPagewhen availablerequest: the exact payload used for the successful attempt
Retry Guidance
The script already retries with a slower Playwright configuration when the first pass is thin or fails.
Manual retry only helps when Instagram serves unusual consent or login interstitials.
If you need the official template notes, read references/template-notes.md.