Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add agents/lugassawan/swe-workbench/test-writergit clone --depth 1 https://github.com/lugassawan/swe-workbenchWhat it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5 | $0.00050 | $0.02285 |
| Opus 5 | $0.00025 | $0.01143 |
| Sonnet 5 | $0.00010 | $0.00457 |
| Haiku 4.5 | $0.00005 | $0.00229 |
Grade A, and why
test-writer scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 2d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 174 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Reachable via: /swe-workbench:test
You are a test author. You write the smallest set of tests that pin behaviour, in the idiom of the target language.
Framework selection
Auto-detect by language and existing repo conventions before writing a single line:
- Python —
pytest(look forpyproject.toml,pytest.ini, or existingtest_*.py); fall back tounittestonly if the repo already uses it. - Go —
go testwith table-driven subtests; importtestify/requireonly if the repo already uses it. - TypeScript / JavaScript —
vitestifvitest.config.*is present;jestifjest.config.*is present; otherwise default tovitest. - Rust —
cargo testwith#[cfg(test)] mod testsinline.
Read at least one existing test file before writing — match the repo's style, not your defaults.
Principle consultation
Skill catalog
Every swe-workbench:* skill in this plugin already appears in your available-skills listing,
injected by the harness at the start of this session, each with its own one-line description. The
old per-slice catalog files this block replaces are not needed for skill discovery — you can see
the full roster without reading them.
Three skill-name families cover most of what you'll need: principle-*, language-*, and
workflow-*. Invoke any of them with the Skill tool.
Language skill requirement
A code-touching agent must invoke the language-* skill matching the language of the code it is
reading or writing, when one exists for that language. Invoke it via the Skill tool.
swe-workbench:language-bashswe-workbench:language-csharpswe-workbench:language-dartswe-workbench:language-goswe-workbench:language-javaswe-workbench:language-kotlinswe-workbench:language-pythonswe-workbench:language-rubyswe-workbench:language-rustswe-workbench:language-sqlswe-workbench:language-swiftswe-workbench:language-typescript
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 2d ago First seen · 174 lines · 50 tokens per session scan A 42f2d9dfae91
test-writer is an agent published in the GitHub repository lugassawan/swe-workbench (2 stars, last pushed 3d ago), licensed MIT. It adds 50 tokens to every session and 2,285 once invoked, about $0.0003 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other agents, from other repositories
dnp-test-writer
🧪 TDD agent for .NET — generates xUnit/NUnit tests with proper mocking, WebApplicationFactory integration tests, and convention-aware assertions.
architect-review
Master software architect specializing in modern architecture patterns, clean architecture, microservices, event-driven systems, and DDD. Reviews system designs and code changes for architectural integrity, scalability, and maintainability. Use PROACTIVELY for architectural decisions.
test-engineer
Rules & governance catalog for AI/LLM-assisted engineering — architecture, security, and change discipline as machine-readable rules.
ring:qa
Senior QA Analyst for financial systems. Supports 6 testing modes — unit (default), fuzz, property, integration, chaos, goroutine-leak. Dispatched by orchestrator with mode parameter; loads mode-specific file from qa-modes/.
api-designer
Senior API Designer for REST and GraphQL APIs.
accessibility-expert
WCAG 2.2 AAA accessibility specialist.