Pro

Technical SEO Audit

Crawl a site politely and get a prioritized fix list for indexation and technical SEO

$59.90

Technical SEO Audit turns your coding agent into a careful technical SEO reviewer. It is for founders, marketers, developers and agencies who need to know why pages are not being crawled or indexed properly, or who are preparing for (or cleaning up after) a site migration.

How it works

A bundled Python crawler (standard library only) walks one host breadth-first at one request per second by default, honours robots.txt and Crawl-delay, and stops at a page cap you set. For every page it records status, redirect chain, title, meta description, H1s, canonical, meta robots and X-Robots-Tag, hreflang, internal and external links, JSON-LD types and more, then compares the crawl with the XML sitemap. It outputs JSON (and optional CSV) with ready-made issue groups such as broken internal links, redirect chains, canonicals pointing at non-200 URLs, noindex pages in the sitemap, orphan candidates and hreflang without return links.

The agent then checks what a crawler cannot: JavaScript rendering, Core Web Vitals (LCP, INP, CLS) using field data you provide, structured data eligibility and international setup, and ranks issues by impact and effort.

What you get

  • SKILL.md workflow with prioritization rules and a chat summary format
  • The crawler and a guide to every issue ID
  • Playbooks: crawl and indexation, Core Web Vitals, structured data and hreflang, internal linking and migrations
  • Audit report template and a CSV issue tracker for developers
  • A worked example of a post-migration audit

The crawler does not render JavaScript. Search engine rules change, so time-sensitive points are marked to verify. No audit can guarantee rankings or traffic.

Highlights

  • Polite stdlib Python crawler: robots-aware, page cap, JSON and CSV output
  • 30+ issue checks: status, redirects, canonicals, noindex, sitemap, hreflang
  • Playbooks for indexation, Core Web Vitals, schema and internal linking
  • Report template and developer issue tracker (CSV)
  • Worked example: a post-migration audit of a SaaS marketing site

Example

Site health check
You

Do a technical SEO audit of https://www.fieldnote.com. We moved domains four months ago and demo requests from Google dropped.

Agent + skill
Technical SEO audit - www.fieldnote.com (400 pages crawled, sample - blog archive not fully crawled)
Health: crawlable and mostly clean, but the domain move left redirect, sitemap and hreflang debris, and 140 integration pages are invisible in the HTML.
Top fixes:
1. [High] Integrations directory links are JS-only - 140 URLs, e.g. /integrations/quickbooks - render <a href> links server-side
2. [High] 312 internal links hit 2-hop redirects - e.g. http://www.fieldnote.com/blog/dispatch-board-tips/ (from /blog/) - link the final https URL
3. [High] Sitemap lists 96 redirecting or 404 URLs - rebuild it from canonical URLs
4. [High] hreflang on /de/ lacks return links (22 sets), plus en-UK should be en-GB
Needs owner data: server logs (optional) | Files: seo-audit/ (report, crawl.json, pages.csv, issue-tracker.csv)
Why is this page not indexed?
You

Google won't index our /pricing page. Can you check what's wrong?

Agent + skill
Checked https://www.example.com/pricing directly and in a 50-page crawl:

- Status 200, no redirect.
- <meta name="robots" content="noindex, follow"> is present in the raw HTML (pricing.tsx line 12 sets robots from the CMS field `hide_from_search`, which is true for this entry).
- Canonical is self-referencing; the page is in the sitemap, so the sitemap and the page send mixed signals (SITEMAP-NOINDEX).
- 38 internal links point to it, depth 1, so discovery is fine.

Fix: untick hide_from_search for the pricing entry, redeploy, then use URL Inspection > Request indexing. Verify the live HTML no longer contains noindex. Indexing timing is up to the search engine.
Staging pre-launch check
You

Before we launch the redesign, crawl our staging site and tell me what will break SEO. It blocks robots.

Agent + skill
Crawled staging with --ignore-robots --max-pages 30 (your own site; robots.txt there is Disallow: /, as expected for staging).

Launch blockers to confirm in production config:
- Every page sends X-Robots-Tag: noindex, nofollow (NOINDEX x30) - make sure this header is staging-only.
- Canonicals are built from the request host (they point at staging.example.com), so production should emit the production host - check the first production deploy before announcing.

Fix before launch:
- 9 internal links to /solutions/* return 404 (HTTP-4XX), e.g. /solutions/dispatch (linked from /, /pricing); the old URLs need 301s to their new equivalents.
- Product template has two H1s and no meta description (H1-MULTIPLE, DESC-MISSING).

Looks good: one canonical per page, valid JSON-LD (Organization, Product, BreadcrumbList), viewport present.

What's inside

technical-seo-audit/
├── agents/
│   └── openai.yaml
├── examples/
│   └── worked-example-b2b-saas-site.md
├── references/
│   ├── core-web-vitals.md
│   ├── crawl-and-indexation.md
│   ├── crawler-issue-guide.md
│   ├── internal-linking-and-architecture.md
│   └── structured-data-and-hreflang.md
├── scripts/
│   └── seo_crawl.py
├── templates/
│   ├── issue-tracker.csv
│   └── technical-seo-audit-report.md
├── LICENSE.txt
├── README.md
└── SKILL.md

Install by unzipping into your agent's skills folder. Install guide →