Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add skills/mehrad-dm/mastermind/roadmapnpx skills add mehrad-dm/mastermind --skill roadmapgit clone --depth 1 https://github.com/mehrad-dm/mastermindWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/mehrad-dm/mastermind/roadmap)<a href="https://agentmods.dev/skills/mehrad-dm/mastermind/roadmap"><img src="https://agentmods.dev/badge/skills/mehrad-dm/mastermind/roadmap.svg" alt="Measured on agentmods" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5 | $0.00088 | $0.01182 |
| Opus 5 | $0.00044 | $0.00591 |
| Sonnet 5 | $0.00018 | $0.00236 |
| Haiku 4.5 | $0.00009 | $0.00118 |
Grade A, and why
roadmap scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 4d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 88 lines — stays where its author put it; the contents beside it link to each section on GitHub.
MasterMind: Map
Long work loses its reasoning long before it loses its code. Six sessions in, the why behind a choice is gone, and the seventh session relitigates it, or silently undoes it. This is the one artifact that outlives the sessions: a decision map the project owns and appends to for the life of the work.
The file: five sections
.mastermind/MAP.md, one per project. Nothing that isn't one of these five belongs in it.
- Destination: what done looks like, in two or three lines. What every decision is measured against.
- Decisions so far: dated, one line each, each with its why. Append-only.
- Open questions: what you can state precisely and cannot answer yet.
- Not yet specified: known unknowns, deliberately unplanned.
- Out of scope: and what it would take to change that.
A decision entry, and the reversal of one, look like this:
2026-03-04 · Sessions live in Postgres, not Redis: one datastore to operate, and
nothing needed sub-millisecond reads. → specs/auth.md
2026-05-19 · REVERSES 2026-03-04: sessions move to Redis. The read path became the
p99 at 340ms once org switching landed. The original reason still held;
the traffic changed. → PR #412
Only section 2 is append-only. The other four describe the present and get rewritten freely: a resolved question leaves "Open questions" the moment it becomes a decision.
Fog of war: don't chart what you can't see
The test for "Open questions" is whether you can state the question precisely today, not whether you can answer it. "Which queue runs the nightly export?" is an open question. "How does billing work?" isn't a question yet; it's a region of fog.
Write the fog as one line naming the unknown and what will lift it: "Billing model, unspecified until the pilot returns pricing feedback." That line is the whole entry. A map carrying a plausible plan for month three reads as decided, and the next session builds against something nobody ever chose.
What ships with it
1 file beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 4d ago First seen · 88 lines · 0 tokens per session scan A 1d6ebfa72177
roadmap is a skill published in the GitHub repository mehrad-dm/mastermind (24 stars, last pushed 4d ago), licensed MIT. It adds 88 tokens to every session and 1,182 once invoked, about $0.0004 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.
Other skills, from other repositories
docker-extend
Use when: User wants to extend Docker with custom tools, personalize the Docker environment, or set up user-specific Docker customization. Triggers: 'extend docker', 'docker-extend', 'add tools to docker', 'customize docker', 'add my tools to the container', 'personalize docker setup', 'docker user setup', 'install…
write-zot-themes
Help the user create, install, or package zot themes, including theme-only extensions.
st-full-workflow
Use when the user asks to run the complete end-to-end Strikethroo workflow for a work order in one shot in this repository — triggers include full workflow, end-to-end, plan and execute, do everything, run the whole strikethroo workflow. Do not use when the user wants only one stage (create a plan, generate tasks, or…
st-code-review
Use when the blueprint execution gate asks for an independent second-harness review of a Strikethroo plan's cumulative diff in this repository — triggers include code review gate, review the plan diff, second-model review, CODEREVIEW hook, review the cumulative diff. Do not use to review a single task, to review code…
st-refine-plan
Use when the user asks to review, refine, improve, interrogate, pressure-test, or update an existing Strikethroo plan by plan ID in this repository — triggers include refine plan, improve plan, review plan, red-team the plan, update plan. Do not use to create a new plan, to generate tasks, or for generic brainstorming…
tlamatini-daily-chat-test
Run the daily automated Tlamatini chat regression — drive a visible Chrome via Playwright, log into agentpage.html, ask up to 1000 curated safe questions one-by-one (Multi-Turn ON, ACPX/Ask-Execs/Exec-Report/Internet OFF), wait for and qualify each answer (heuristic + LLM judge on failures), then write a dated report…