seo-sitemap

seo-sitemap is a skill for Claude Code from amirjahfar1/automate-seo-with-claude. It costs 139 tokens per session (2,963 once invoked), scanned A, original, MIT.

A tool for comparing the URLs listed in a website's XML sitemap with pages that a crawler can find and Google has indexed. An XML sitemap is a file that tells search engines which pages a site considers important; GSC is Google Search Console.

In plain words
What is it for?
Use it to audit sitemap coverage, find orphaned pages, check Google sitemap processing, and compare crawling and indexing results for a domain.
Why use it?
It reveals pages listed in the sitemap that cannot be reached, as well as reachable or indexed pages missing from the sitemap.

Skill for Claude Code

Written for Claude Code: shipped in a Claude Code plugin.

Part of the seo-skills plugin — 26 skills shipped together

Good fit Use it to audit sitemap coverage, find orphaned pages, check Google sitemap processing, and compare crawling and indexing results for a domain.

Compare 6 skills from other repositories ↓
Install with agentmods
npx agentmods add skills/amirjahfar1/automate-seo-with-claude/seo-sitemap
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Any agent
npx skills add amirjahfar1/automate-seo-with-claude --skill seo-sitemap
Clone the repo
git clone --depth 1 https://github.com/amirjahfar1/automate-seo-with-claude

Made for: Claude Code.

Or install seo-skills, the plugin that ships this one along with the rest of its 26 skills.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for seo-sitemap

README.md
[![agentmods](https://agentmods.dev/badge/skills/amirjahfar1/automate-seo-with-claude/seo-sitemap/github.svg)](https://agentmods.dev/skills/amirjahfar1/automate-seo-with-claude/seo-sitemap)
Your own site
<a href="https://agentmods.dev/skills/amirjahfar1/automate-seo-with-claude/seo-sitemap"><img src="https://agentmods.dev/badge/skills/amirjahfar1/automate-seo-with-claude/seo-sitemap/github.svg" alt="Measured on agentmods" height="20"></a>

Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.

agentmods 80×15 button for seo-sitemap

Your own site · 80×15
<a href="https://agentmods.dev/skills/amirjahfar1/automate-seo-with-claude/seo-sitemap"><img src="https://agentmods.dev/badge/skills/amirjahfar1/automate-seo-with-claude/seo-sitemap.svg" alt="Reviewed on agentmods" width="80" height="20"></a>
Per session 139 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 2,963 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00139 $0.02963
Opus 5 $0.00069 $0.01482
Sonnet 5 $0.00028 $0.00593
Haiku 4.5 $0.00014 $0.00296

Measured 12d ago against content hash 4e409b02ab00, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-12, from the pricing page.

Security

Grade A, and why

seo-sitemap scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 12d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

skills/seo-sitemap/SKILL.md · 141 lines

How it starts

The opening of the file, as written. The whole thing — 141 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Example output: examples/seo-sitemap-notion-so-20260514/SITEMAP.md

Sitemap Analysis

Compare a domain's XML sitemap against what's actually crawled and indexed — GSC indexed pages (get_search_analytics dimensions=["page"]), GSC sitemap ingestion (get_sitemap_details), and a DataForSEO on_page_instant_pages fetch loop over the declared/indexed URL set. Surface what the sitemap claims vs what's really reachable, in both directions.

Prerequisites

  • DataForSEO MCP server connected.
  • GSC (mcp__gscServer__*) recommended — it's the authoritative source for which pages Google indexes and which sitemaps Google ingested (the comparison baseline). Firecrawl optional — for URL discovery (firecrawl_map) when the sitemap is missing or suspect.
  • Claude's WebFetch tool available.
  • User provides: a target domain. Optional: the sitemap URL if not at /sitemap.xml (auto-discovery from robots.txt is attempted first).
  • Predecessor (recommended): seo-technical-audit on this domain — its discovered URL set + On-Page fetch results can be reused as the crawl baseline. Without it, this skill builds its own baseline from GSC indexed pages + an On-Page fetch loop.

Optional accelerator: if you have a hosted-crawl MCP you can substitute it for the URL-discovery + fetch loop — not required.

Process

  1. Validate target & build the crawl baseline
    • Normalise the domain.
    • Build the "what's really crawled/indexed" baseline this skill compares the sitemap against:
      • GSC indexed pagesmcp__gscServer__get_search_analytics with dimensions=["page"] (pages Google indexes, with clicks/impressions). Authoritative for the user's own verified property.
      • DataForSEO On-Page fetch loopmcp__dataforseo__on_page_instant_pages over the discovered URL set (declared sitemap URLs + GSC indexed pages), capped to a top-N ceiling — gives status code, redirects, indexability, depth signals per page.
      • DataForSEO top pagesmcp__dataforseo__dataforseo_labs_google_relevant_pages (broader domain page inventory than the sitemap in some cases).
    • If seo-technical-audit already ran on this domain, reuse its discovered URL set + On-Page results as the baseline instead of re-fetching.
    • Firecrawl availability check. If mcp__firecrawl-mcp__firecrawl_map is available, Mode-2 (URL discovery via crawl) is offered when the sitemap is missing or suspect. Cost: ~0.5 Firecrawl credits per URL discovered, hard cap 500 URLs (~250 credits). Without Firecrawl, the skill runs Mode-1 only and notes the gap if Mode-2 was needed. User may pass --no-firecrawl to force Mode-1 even when Firecrawl is available (saves credits at the cost of orphan/missing analysis when sitemap is broken).

Read the full file on GitHub · 141 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 12d ago First seen · 141 lines · 139 tokens per session scan A 4e409b02ab00

Subscribe to this mod's changes

seo-sitemap is a skill published in the GitHub repository amirjahfar1/automate-seo-with-claude (2 stars, last pushed 3mo ago), licensed MIT. It adds 139 tokens to every session and 2,963 once invoked, about $0.0007 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other skills, from other repositories

seo-site-audit-pro

Flagship comprehensive SEO audit combining Ahrefs and GSC data in sequential waves with checkpoint saves. Use when user says "site audit pro", "full audit", "comprehensive audit", "audit with live data", "deep audit", "pro audit", "complete SEO audit", "run a full site audit", or "audit everything for this domain".…

lionkiii/claude-seo-skills · 90 tokens

seo-content

Content quality and E-E-A-T analysis with AI citation readiness assessment. Enhanced with live Ahrefs (actual keyword rankings, positions) and GSC (search query performance) data to validate static E-E-A-T analysis with real user behavior. Use when user says "content quality", "E-E-A-T", "content analysis"…

lionkiii/claude-seo-skills · 83 tokens

seo-core-web-vitals

Dedicated Core Web Vitals deep-dive: pull field data (CrUX / PageSpeed Insights), fall back to Lighthouse lab data, score LCP, INP, and CLS against Google's thresholds at the 75th percentile, and run per-metric diagnosis playbooks with concrete fixes. Use when user says "core web vitals", "CWV", "LCP", "INP", "CLS"…

lionkiii/claude-seo-skills · 99 tokens

seo-geo

Optimize content for AI Overviews (formerly SGE), ChatGPT web search, Perplexity, and other AI-powered search experiences. GEO analysis enriched with Ahrefs Brand Radar AI visibility data when available. Includes brand mention signals, AI crawler accessibility, llms.txt compliance, passage-level citability scoring…

lionkiii/claude-seo-skills · 119 tokens

seo-lighthouse-audit

Run a Lighthouse audit against any URL via the local Lighthouse CLI or the PageSpeed Insights API fallback, parse category scores and failed audits, and map every failed audit ID to its official Chrome docs page plus a one-line fix. Use when user says "lighthouse", "lighthouse audit", "run lighthouse", "lighthouse…

lionkiii/claude-seo-skills · 78 tokens

seo-llms-txt

Generate, validate, or audit llms.txt files for AI search visibility. Crawls site structure, generates spec-compliant Markdown index for LLMs. Use when user says "llms.txt", "llm txt", "AI crawlers", "generate llms", "LLM file", "AI readability file".

lionkiii/claude-seo-skills · 72 tokens