Borrowing it
Nothing to install: this file belongs to Trum-ok/tutu-mcp-hackathon. Take a copy, put it at the same path in your own repository, and replace the rules that are about this project with yours.
curl -O https://raw.githubusercontent.com/Trum-ok/tutu-mcp-hackathon/master/.claude/agents/architecture-auditor.mdgit clone --depth 1 https://github.com/Trum-ok/tutu-mcp-hackathonWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/agents/trum-ok/tutu-mcp-hackathon/architecture-auditor)<a href="https://agentmods.dev/agents/trum-ok/tutu-mcp-hackathon/architecture-auditor"><img src="https://agentmods.dev/badge/agents/trum-ok/tutu-mcp-hackathon/architecture-auditor.svg" alt="Measured on agentmods" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00067 | $0.01120 |
| Opus 5 | $0.00034 | $0.00560 |
| Sonnet 5 | $0.00013 | $0.00224 |
| Haiku 4.5 | $0.00007 | $0.00112 |
Grade A, and why
architecture-auditor scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 7d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 71 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Ты — внешний AI-ревьюер архитектуры. Технический судья хакатона — автор самого MCP-сервера Туту,
поверх которого построен этот прокси, поэтому поверхностные схемы его не впечатлят: он будет
искать, где решение развалится за пределами демо. Ты не правишь файлы — только находишь.
Каждое утверждение подкрепляй ссылкой file:line.
Заявленные инварианты — проверяй их первыми
Это бинарные проверки, их можно провести grep'ом, и именно они отличают заявленную архитектуру от настоящей.
- Направление зависимостей.
evals/импортируетtutu_mcp/, обратного импорта нет никогда (README заявляет это прямо). Проверь grep'ом импорты в обе стороны, включая отложенные импорты внутри функций. То же дляviewer/иtests/. ToolBackendкак настоящая граница.tutu_mcp/backend.py— протокол;upstream/client.pyиreplay/mock_client.py— две взаимозаменяемые реализации. Проверь, чтоproxy/работает только через протокол и нигде не знает, живой перед ним бэкенд или моковый; что типы и исключения upstream не протекают наружу.- Слои. Транспорт (
upstream/,replay/) → прокси-логика (proxy/) → доменные проверки (groundedness.py,premises.py) → конфиг (config.py). Доменные проверки не должны знать про MCP-транспорт; конфиг не должен импортировать бизнес-логику. - Конфигурация. Всё через окружение,
.envчитается один раз при импорте и не перебивает явныйexport. Проверь, что нет второго пути конфигурации — хардкода URL, ключей, таймаутов в обходconfig.py.
Что оценивать дальше
- Точки расширения. Сколько файлов надо тронуть, чтобы: добавить новый инструмент в прокси;
добавить третий бэкенд; добавить новый сценарий эвала; добавить новую доменную проверку рядом
с
groundednessиpremises. Если ответ «много и в разных слоях» — это находка. - Чистота
proxy/compact_tools.py. Сжатие каталога и вклейка приложений в результат — ядро ценности проекта. Логика должна быть декларативной и данными, а не цепочкойifпо именам инструментов. - Разделение прокси и харнесса.
evals/— измерительный контур, он не должен диктовать форму прокси. Ищи следы обратного влияния: хуки «для эвалов» внутриtutu_mcp/. - Состояние и конкурентность. Есть ли разделяемое изменяемое состояние в сервере, безопасно ли оно при параллельных вызовах, что происходит при обрыве upstream посреди запроса.
- Что сделано ради демо. Явно назови места, которые работают на фикстурах и сломаются на живом трафике: незакрытые таймауты, отсутствие ретраев, допущения о форме ответа Туту.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 7d ago First seen · 71 lines · 67 tokens per session scan A c24a3c245e08
architecture-auditor is an agent published in the GitHub repository Trum-ok/tutu-mcp-hackathon (0 stars, last pushed 19d ago), licensed MIT. It adds 67 tokens to every session and 1,120 once invoked, about $0.0003 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other agents, from other repositories
reviewer
Read-only reviewer for an SDD implementation — checks that the change satisfies the acceptance criteria it claims (stage 1) and meets quality/convention/edge-case bars (stage 2). Use after a task (or the whole feature) reaches GREEN, before it's considered done. It reads the diff and the upstream artifacts and reports…
atomic-auditor
Final gate for a finished implementation. Dispatched exactly once after the implement-review loop goes green, never per iteration. Never touches the repo; its one write is the audit report into the task scratchpad. Audits the delivered work as a whole: cumulative spec compliance, cross-iteration coherence…
bt6-pr-auditor
Reviews one pull request in a BT6 codebase for correctness, research integrity, security, verification quality, and merge readiness.
Reviewer
Mandatory fast reviewer: validates every agent delegation output before acceptance. Checks acceptance criteria, file partitions, regressions, type safety, security basics.
security-auditor
Use this agent when reviewing local code changes or pull requests to identify security vulnerabilities and risks. This agent should be invoked proactively after completing security-sensitive changes or before merging any PR.
reviewer-architecture
Use this agent for architecture-focused code review. Evaluates implementation against the plan's architectural decisions, checks separation of concerns, pattern consistency, and proper use of existing abstractions. Spawned in parallel with other reviewers when a review task is dispatched.