Full Website SEO Audit
Process
- Fetch homepage: use
scripts/fetch_page.py to retrieve HTML
- Detect business type: analyze homepage signals per seo orchestrator
- Crawl site: follow internal links up to 500 pages, respect robots.txt
- Delegate to subagents (if available, otherwise run inline sequentially):
seo-technical -- robots.txt, sitemaps, canonicals, Core Web Vitals, security headers
seo-content -- E-E-A-T, readability, thin content, AI citation readiness
seo-schema -- detection, validation, generation recommendations
seo-sitemap -- structure analysis, quality gates, missing pages
seo-performance -- LCP, INP, CLS measurements
seo-visual -- screenshots, mobile testing, above-fold analysis
seo-geo -- AI crawler access, llms.txt, citability, brand mention signals
seo-local -- GBP signals, NAP consistency, reviews, local schema, industry-specific local factors (spawn when Local Service industry detected: brick-and-mortar, SAB, or hybrid business type)
seo-maps -- Geo-grid rank tracking, GBP audit, review intelligence, competitor radius mapping (spawn when Local Service detected AND DataForSEO MCP available)
seo-google -- CWV field data (CrUX), URL indexation (GSC), organic traffic (GA4) (spawn when Google API credentials detected via python scripts/google_auth.py --check)
seo-backlinks -- Backlink profile data: DA/PA, referring domains, anchor text, toxic links (spawn when Moz or Bing API credentials detected via python scripts/backlinks_auth.py --check, or always include Common Crawl domain-level metrics)
- Score -- aggregate into SEO Health Score (0-100)
- Report -- generate prioritized action plan
Crawl Configuration
Max pages: 500
Respect robots.txt: Yes
Follow redirects: Yes (max 3 hops)
Timeout per page: 30 seconds
Concurrent requests: 5
Delay between requests: 1 second
Output Files
FULL-AUDIT-REPORT.md: Comprehensive findings
ACTION-PLAN.md: Prioritized recommendations (Critical > High > Medium > Low)
screenshots/: Desktop + mobile captures (if Playwright available)
- PDF Report (recommended): Generate a professional A4 PDF using
scripts/google_report.py --type full. This produces a white-cover enterprise report with TOC, executive summary, charts (Lighthouse gauges, query bars, index donut), metric cards, threshold tables, prioritized recommendations with effort estimates, and implementation roadmap. Always offer PDF generation after completing an audit.
Scoring Weights
| Category |
Weight |
| Technical SEO |
22% |
| Content Quality |
23% |
| On-Page SEO |
20% |
| Schema / Structured Data |
10% |
| Performance (CWV) |
10% |
| AI Search Readiness |
10% |
| Images |
5% |
Report Structure
Executive Summary
- Overall SEO Health Score (0-100)
- Business type detected
- Top 5 critical issues
- Top 5 quick wins
Technical SEO
- Crawlability issues
- Indexability problems
- Security concerns
- Core Web Vitals status
Content Quality
- E-E-A-T assessment
- Thin content pages
- Duplicate content issues
- Readability scores
On-Page SEO
- Title tag issues
- Meta description problems
- Heading structure
- Internal linking gaps
Schema & Structured Data
- Current implementation
- Validation errors
- Missing opportunities
Performance
- LCP, INP, CLS scores
- Resource optimization needs
- Third-party script impact
Images
- Missing alt text
- Oversized images
- Format recommendations
AI Search Readiness
- Citability score
- Structural improvements
- Authority signals
Priority Definitions
- Critical: Blocks indexing or causes penalties (fix immediately)
- High: Significantly impacts rankings (fix within 1 week)
- Medium: Optimization opportunity (fix within 1 month)
- Low: Nice to have (backlog)
DataForSEO Integration (Optional)
If DataForSEO MCP tools are available, spawn the seo-dataforseo agent alongside existing subagents to enrich the audit with live data: real SERP positions, backlink profiles with spam scores, on-page analysis (Lighthouse), business listings, and AI visibility checks (ChatGPT scraper, LLM mentions).
Google API Integration (Optional)
If Google API credentials are configured (python scripts/google_auth.py --check), spawn the seo-google agent to enrich the audit with real Google field data: CrUX Core Web Vitals (replaces lab-only estimates), GSC URL indexation status, search performance (clicks, impressions, CTR), and GA4 organic traffic trends. The Performance (CWV) category score benefits most from field data.
Error Handling
| Scenario |
Action |
| URL unreachable (DNS failure, connection refused) |
Report the error clearly. Do not guess site content. Suggest the user verify the URL and try again. |
| robots.txt blocks crawling |
Report which paths are blocked. Analyze only accessible pages and note the limitation in the report. |
| Rate limiting (429 responses) |
Back off and reduce concurrent requests. Report partial results with a note on which sections could not be completed. |
| Timeout on large sites (500+ pages) |
Cap the crawl at the timeout limit. Report findings for pages crawled and estimate total site scope. |
1---2name: seo-audit3description: Full website SEO audit with parallel subagent delegation. Crawls up to 500 pages, detects business type, delegates to 10 specialists (7 core + 3 conditional), generates health score. Use when user says audit, full SEO check, analyze my site, or website health check.4license: MIT5---67# Full Website SEO Audit89## Process10111. **Fetch homepage**: use `scripts/fetch_page.py` to retrieve HTML122. **Detect business type**: analyze homepage signals per seo orchestrator133. **Crawl site**: follow internal links up to 500 pages, respect robots.txt144. **Delegate to subagents** (if available, otherwise run inline sequentially):15 - `seo-technical` -- robots.txt, sitemaps, canonicals, Core Web Vitals, security headers16 - `seo-content` -- E-E-A-T, readability, thin content, AI citation readiness17 - `seo-schema` -- detection, validation, generation recommendations18 - `seo-sitemap` -- structure analysis, quality gates, missing pages19 - `seo-performance` -- LCP, INP, CLS measurements20 - `seo-visual` -- screenshots, mobile testing, above-fold analysis21 - `seo-geo` -- AI crawler access, llms.txt, citability, brand mention signals22 - `seo-local` -- GBP signals, NAP consistency, reviews, local schema, industry-specific local factors (spawn when Local Service industry detected: brick-and-mortar, SAB, or hybrid business type)23 - `seo-maps` -- Geo-grid rank tracking, GBP audit, review intelligence, competitor radius mapping (spawn when Local Service detected AND DataForSEO MCP available)24 - `seo-google` -- CWV field data (CrUX), URL indexation (GSC), organic traffic (GA4) (spawn when Google API credentials detected via `python scripts/google_auth.py --check`)25 - `seo-backlinks` -- Backlink profile data: DA/PA, referring domains, anchor text, toxic links (spawn when Moz or Bing API credentials detected via `python scripts/backlinks_auth.py --check`, or always include Common Crawl domain-level metrics)265. **Score** -- aggregate into SEO Health Score (0-100)276. **Report** -- generate prioritized action plan2829## Crawl Configuration3031```32Max pages: 50033Respect robots.txt: Yes34Follow redirects: Yes (max 3 hops)35Timeout per page: 30 seconds36Concurrent requests: 537Delay between requests: 1 second38```3940## Output Files4142- `FULL-AUDIT-REPORT.md`: Comprehensive findings43- `ACTION-PLAN.md`: Prioritized recommendations (Critical > High > Medium > Low)44- `screenshots/`: Desktop + mobile captures (if Playwright available)45- **PDF Report** (recommended): Generate a professional A4 PDF using `scripts/google_report.py --type full`. This produces a white-cover enterprise report with TOC, executive summary, charts (Lighthouse gauges, query bars, index donut), metric cards, threshold tables, prioritized recommendations with effort estimates, and implementation roadmap. Always offer PDF generation after completing an audit.4647## Scoring Weights4849| Category | Weight |50|----------|--------|51| Technical SEO | 22% |52| Content Quality | 23% |53| On-Page SEO | 20% |54| Schema / Structured Data | 10% |55| Performance (CWV) | 10% |56| AI Search Readiness | 10% |57| Images | 5% |5859## Report Structure6061### Executive Summary62- Overall SEO Health Score (0-100)63- Business type detected64- Top 5 critical issues65- Top 5 quick wins6667### Technical SEO68- Crawlability issues69- Indexability problems70- Security concerns71- Core Web Vitals status7273### Content Quality74- E-E-A-T assessment75- Thin content pages76- Duplicate content issues77- Readability scores7879### On-Page SEO80- Title tag issues81- Meta description problems82- Heading structure83- Internal linking gaps8485### Schema & Structured Data86- Current implementation87- Validation errors88- Missing opportunities8990### Performance91- LCP, INP, CLS scores92- Resource optimization needs93- Third-party script impact9495### Images96- Missing alt text97- Oversized images98- Format recommendations99100### AI Search Readiness101- Citability score102- Structural improvements103- Authority signals104105## Priority Definitions106107- **Critical**: Blocks indexing or causes penalties (fix immediately)108- **High**: Significantly impacts rankings (fix within 1 week)109- **Medium**: Optimization opportunity (fix within 1 month)110- **Low**: Nice to have (backlog)111112## DataForSEO Integration (Optional)113114If DataForSEO MCP tools are available, spawn the `seo-dataforseo` agent alongside existing subagents to enrich the audit with live data: real SERP positions, backlink profiles with spam scores, on-page analysis (Lighthouse), business listings, and AI visibility checks (ChatGPT scraper, LLM mentions).115116## Google API Integration (Optional)117118If Google API credentials are configured (`python scripts/google_auth.py --check`), spawn the `seo-google` agent to enrich the audit with real Google field data: CrUX Core Web Vitals (replaces lab-only estimates), GSC URL indexation status, search performance (clicks, impressions, CTR), and GA4 organic traffic trends. The Performance (CWV) category score benefits most from field data.119120## Error Handling121122| Scenario | Action |123|----------|--------|124| URL unreachable (DNS failure, connection refused) | Report the error clearly. Do not guess site content. Suggest the user verify the URL and try again. |125| robots.txt blocks crawling | Report which paths are blocked. Analyze only accessible pages and note the limitation in the report. |126| Rate limiting (429 responses) | Back off and reduce concurrent requests. Report partial results with a note on which sections could not be completed. |127| Timeout on large sites (500+ pages) | Cap the crawl at the timeout limit. Report findings for pages crawled and estimate total site scope. |