zytedata
- 15 skills
- 0 followers
- 6 hours ago last updated
- ▌ Scrape · zytedata bundleEnd-to-end web scraping workflow — from URL to working spider with web-poet page objects. Use this for full-site or multiple-page crawls that create a new scraping workflow. Do not use for fixing, debugging, or modifying an existing spider.
- ▌ Scrape Spec · zytedataExpand a spec created by /scrape-define — download more pages, compare HTML variants, extract values, optional browser review
- ▌ Scrape Define · zytedataCreate a new extraction spec from a URL — explore a detail page, discover fields, quick schema approval
- ▌
- ▌ Scrape Zyte Login · zytedata bundleSet up the user's Zyte account and credentials. Use when the user asks to set up, log in, sign up, or get an API key for Zyte; when ZYTE_API_KEY is missing; when a site is blocked; or before Scrapy Cloud deployment.
- ▌ Scrape Analyze Page · zytedata bundleExtract structured data (all available fields with values) from a page saved locally as an HTML file, optionally following a schema. Use this skill only to process already downloaded files. Do not invoke when the user provides a URL. When invoking, pass the user's full request verbatim as args — do not pre-parse file paths and don't rephrase it.
- ▌ Scrape Explore Site · zytedata bundleExplore a website to find and save diverse pages (start, list, detail) with classified links
- ▌ Scrape Scrapy Cloud · zytedata bundleGeneral-purpose Scrapy Cloud skill — deploy projects, schedule spiders, list/stop jobs, and view items or logs. Use when asked to deploy a project or spider to Scrapy Cloud / Zyte Cloud / Scrapinghub, schedule or run a spider remotely, manage jobs, or inspect scraped items and logs.
- ▌
- ▌ Scrape Review Schema · zytedata bundleGenerate an HTML review page for schema and extracted data verification
- ▌ Scrape Ensure Project · zytedata bundleEnsure a Scrapy project exists with scrapy-poet and Zyte API support
- ▌ Scrape Zyte API Stats · zytedata bundleQuery usage stats for Zyte API requests that have already been sent, including cost, request volume, response times, status codes, filters, grouping, and pagination. Use this skill whenever answering the request requires looking up recorded Zyte API usage, spend, request counts, response times, grouped stats, or filtered stats — including projecting future spend or usage when the projection explicitly extrapolates from actual recorded usage (e.g. "based on this month's usage so far, project month-end spend"). Do not use it when the user asks what a spider, job, or workload might, would, or could generate — those are hypothetical estimates with no recorded data to query, even if a spider exists in the project. For a projection to qualify, the user must explicitly reference past or current recorded usage as the basis. Also do not use for dashboard setup, documentation, configuration, or account-help questions.
- ▌
- ▌ Scrape Codegen Analyze · zytedataAnalyze an HTML page to produce field extraction instructions for code generation
- ▌ Scrape Codegen Generate · zytedataGenerate web-poet page object code from per-page extraction analyses