Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
git clone --depth 1 https://github.com/navraj007in/architecture-cowork-pluginWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/commands/navraj007in/architecture-cowork-plugin/optimize-costs)<a href="https://agentmods.dev/commands/navraj007in/architecture-cowork-plugin/optimize-costs"><img src="https://agentmods.dev/badge/commands/navraj007in/architecture-cowork-plugin/optimize-costs/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/commands/navraj007in/architecture-cowork-plugin/optimize-costs"><img src="https://agentmods.dev/badge/commands/navraj007in/architecture-cowork-plugin/optimize-costs.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00010 | $0.03861 |
| Opus 5 | $0.00005 | $0.01930 |
| Sonnet 5 | $0.00002 | $0.00772 |
| Haiku 4.5 | $0.00001 | $0.00386 |
Grade A, and why
optimize-costs scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 11d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 453 lines — stays where its author put it; the contents beside it link to each section on GitHub.
/architect:optimize-costs
Analyzes current infrastructure costs and recommends 5-10 savings opportunities. For each saving, shows: monthly savings, trade-off, effort to implement, and ROI.
Instead of "cut costs anywhere", users get: "Here are 8 ways to save $3k/month. Ranked by savings. Pick which you're comfortable with."
Trigger
/architect:optimize-costs
/architect:optimize-costs [--target-monthly 3000] # hit specific budget
/architect:optimize-costs [--target-percent -20] # reduce costs by 20%
/architect:optimize-costs [--exclude database] # don't touch certain services
/architect:optimize-costs [--show-all] # show all 20+ possible optimizations
Purpose
Infrastructure costs creep upward:
- Database grows, per-query cost increases
- Traffic spikes, compute needs double
- Monitoring gets more verbose (logs cost more)
- You suddenly realize you're paying $15k/month
This command stops the creep. It finds low-hanging fruit ($500-5k/month savings) without hurting product.
Input
Context loading: Read architecture-output/_state.json.cost_estimate first if it exists — fall back to architecture-output/cost-estimate.md for detail not in _state.json.
If neither exists and the user has not supplied current spend interactively: "I need a cost baseline to optimize from. Run
/architect:cost-estimatefirst, then come back here."
Required: Current cost estimate or actual spend
{
"current_monthly_cost": 8500,
"target_monthly_cost": 5500, // optional: hit this budget
"current_dau": 100000, // expected users
"scaling_plan_months": 6, // when expect to double users?
"can_sacrifice": {
"latency": false, // can we be slower?
"uptime": false, // can we have more downtime?
"features": false, // can we cut features?
"scale": true // can we handle less traffic?
}
}
Output
1. architecture-output/cost-optimization-plan.md
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 11d ago First seen · 453 lines · 10 tokens per session scan A 412da8af2792
optimize-costs is a command published in the GitHub repository navraj007in/architecture-cowork-plugin (2 stars, last pushed 2mo ago), licensed Apache-2.0. It adds 10 tokens to every session and 3,861 once invoked, about $0.0001 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other commands, from other repositories
nyann:doctor
Run nyann's read-only hygiene + documentation audit on the current repo. Emits the same drift sections retrofit uses, but never offers remediation or writes to the filesystem. Use this on session start, in CI, or any time you want to sanity-check a repo without risking mutations.
nyann:retrofit
Audit an existing repo against a profile and fix what's drifted. Unlike doctor (read-only), retrofit detects missing hooks, misconfigured gitignore, documentation gaps, and non-compliant history, then offers to remediate via bootstrap. Idempotent — safe to re-run.
nyann:apply
Apply an Infrastructure-as-Code change — the highest-stakes mutator in nyann; it can change real cloud infrastructure. Re-runs the plan, shows it, confirms, then applies. Unmistakably opt-in: apply is never the default and destructive applies require a second explicit confirm. For IaC apply intent only (not "apply a…
nyann:hotfix
Create the branch topology for a patch release against a previously tagged version. Ensures release/ . exists from the source tag, then creates hotfix/ off it. After this, the user commits the fix and runs /nyann:release from the hotfix branch.
nyann:release
Cut a release: generate a CHANGELOG section from Conventional Commits, make a release commit, and create an annotated tag. Defaults to conventional-changelog strategy.
nyann:ship
Open a GitHub pull request AND merge it in one step. Default uses GitHub's native auto-merge (returns immediately with outcome:"queued"); --client-side polls for green CI in the foreground then runs gh pr merge. Requires gh installed + authed.