registry-crawler

registry-crawler is an agent for coding agents from ea-toolkit/architecture-catalog. It costs 65 tokens per session (464 once invoked), scanned A, original, MIT.

An agent that reads existing architecture documents—such as wiki exports, spreadsheets, text files, or diagrams—and proposes entries for an architecture registry.

In plain words
What is it for?
Use it to import systems, services, APIs, data objects, capabilities, and other architecture elements, review proposed entries, create approved registry files, and validate them.
Why use it?
It turns scattered documentation into a consistent catalogue while marking uncertain information as TBD instead of guessing.

Agent

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add agents/ea-toolkit/architecture-catalog/registry-crawler
Clone the repo
git clone --depth 1 https://github.com/ea-toolkit/architecture-catalog

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for registry-crawler

README.md
[![agentmods](https://agentmods.dev/badge/agents/ea-toolkit/architecture-catalog/registry-crawler.svg)](https://agentmods.dev/agents/ea-toolkit/architecture-catalog/registry-crawler)
Your own site
<a href="https://agentmods.dev/agents/ea-toolkit/architecture-catalog/registry-crawler"><img src="https://agentmods.dev/badge/agents/ea-toolkit/architecture-catalog/registry-crawler.svg" alt="Measured on agentmods" height="20"></a>
Per session 65 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 464 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00065 $0.00464
Opus 5 $0.00032 $0.00232
Sonnet 5 $0.00013 $0.00093
Haiku 4.5 $0.00006 $0.00046

Measured 4d ago against content hash 1b252635397d, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

registry-crawler scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 4d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

templates/agents/registry-crawler.md · 47 lines

What it actually says

Registry Crawler — Documentation Importer

You help onboard existing architecture documentation into the catalog by reading documents and proposing structured registry entries.

Process

  1. Discover the schema: Read models/registry-mapping.yaml to understand:

    • Available element types (under elements:) — each with label, layer, folder, fields
    • Available layers (under layers:)
    • Relationship types (under relationships:)
  2. Read source documents provided by the user (wiki export, markdown files, text files, CSV).

  3. Identify architecture elements — systems, services, APIs, data objects, capabilities, etc.

  4. For each identified element:

    • Determine the best element type from registry-mapping.yaml
    • Read the _template.md for that type (found in the type's folder path)
    • Draft a registry entry with YAML frontmatter + description
    • Use TBD for uncertain fields — never guess
  5. Present all proposed entries in a table for user review:

    # Name Type Layer Domain Confidence
  6. On approval, create the files in registry-v2/ in the correct sub-folders.

  7. Run validation: python scripts/validate.py to verify.

Rules

  • Never auto-create without user review — always present proposals first
  • One element per file, kebab-case file names
  • Mark unknown fields as TBD rather than guessing
  • Err on the side of fewer, well-defined elements over many vague ones
  • Check for duplicates before proposing (search existing entries)
  • Discover folder paths from registry-mapping.yaml — never hardcode them
Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 4d ago First seen · 47 lines · 65 tokens per session scan A 1b252635397d

Subscribe to this mod's changes

registry-crawler is an agent published in the GitHub repository ea-toolkit/architecture-catalog (38 stars, last pushed 4mo ago), licensed MIT. It adds 65 tokens to every session and 464 once invoked, about $0.0003 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other agents, from other repositories

auto-reviewer

Automated PR reviewer for the EmDash CI workflows. Used by the /review and /ultrareview workflows to leave structured review feedback on pull requests. Not intended for interactive local use -- prefer the default build or plan agents for that.

emdash-cms/emdash · 53 tokens

discovery

How an agent finds these surfaces without being told: the API catalog, DNS-based discovery, an agent-skills index, and the places people actually look.

DuvInc/duvlify · 31 tokens

overview

Duvlify defines six agent surfaces and four tools once, then adapts them to MCP, plain HTTP and WebMCP so they always agree.

DuvInc/duvlify · 30 tokens

mcp

The page outline agents read for line offsets, the in-browser WebMCP bridge and when to run it yourself, the rate limits, and how the surfaces are tested.

DuvInc/duvlify · 34 tokens

prompt-debugger

Evaluates why a prompt produced bad, unexpected, or suboptimal output and suggests targeted fixes. Use when a user says "my prompt isn't working", "this prompt gives bad results", "why is my prompt failing", "debug this prompt", "the AI keeps getting this wrong", "fix my prompt", "prompt not producing expected…

RadOrigin-LLC/RAD-Claude-Skills · 456 tokens

lp-brandmark

Brandmark designer (LandingForge phase 3.5 — after lp-designer's tokens/design, before/with lp-builder). Turns the committed visual system into a single BOLD, SOLID, FLAT logo mark and the full favicon/icon/OG asset set. Reads design.md (accent hex, mood, light|dark base) + brief.md (product, positioning → mark…

NmadeleiDev/landingforge · 212 tokens