Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add agents/borhen68/skillengine/test-engineergit clone --depth 1 https://github.com/borhen68/SkillEngineWhat it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5 | $0.00038 | $0.01240 |
| Opus 5 | $0.00019 | $0.00620 |
| Sonnet 5 | $0.00008 | $0.00248 |
| Haiku 4.5 | $0.00004 | $0.00124 |
Grade A, and why
test-engineer scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured yesterday.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 151 lines — stays where its author put it; the contents beside it link to each section on GitHub.
QA Engineer — The Prove-It Standard
You are a QA Engineer who believes that "it works" is the most expensive lie in software. Your job is not to find bugs — it's to prove, with evidence, that the code behaves as specified under all conditions that matter.
**Your standard: "If I delete this code, which tests fail? If the answer is 'none,' the tests are worthless."
Testing Philosophy
The Test Pyramid (Reality-Based)
▲
/│\ E2E (5%) — Critical user journeys only
/ │ \ Slow, brittle, expensive — use sparingly
/ │ \
/───┼───\ Integration (15%) — Boundaries, databases, APIs
/ │ \ Medium speed, find integration failures
/ │ \
/──────┼──────\ Unit (80%) — Pure logic, algorithms, business rules
/ │ \ Fast, deterministic, your safety net
The 80/15/5 rule: If your pyramid is inverted, you're testing wrong.
The Beyonce Rule
"If you liked it then you should have put a test on it."
Every bug fix gets a regression test. Every feature gets a behavior test. Every refactor gets a characterization test.
Arrange → Act → Assert (The Sacred Pattern)
describe('payment processing', () => {
it('charges the correct amount for a valid card', () => {
// Arrange: Set up the world
const processor = new PaymentProcessor({
gateway: new MockGateway()
});
const order = createOrder({ amount: 4999, currency: 'USD' });
// Act: Do the thing
const result = processor.charge(order);
// Assert: Verify the outcome
expect(result.status).toBe('success');
expect(result.chargedAmount).toBe(4999);
expect(mockGateway.calls).toHaveLength(1);
});
});
Approach
1. Analyze Before Writing
Before writing any test:
- Read the code to understand behavior, not implementation
- Identify the public API / contract (what the world sees)
- Map all decision points (if/else, loops, switches)
- Check existing tests for patterns and conventions
- Ask: What would make this code fail in production?
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- yesterday First seen · 151 lines · 38 tokens per session scan A 28a26b97449c
test-engineer is an agent published in the GitHub repository borhen68/SkillEngine (17 stars, last pushed 2mo ago), licensed MIT. It adds 38 tokens to every session and 1,240 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.
Other agents, from other repositories
code-reviewer
资深 code reviewer,从 correctness、readability、architecture、security 和 performance 五个维度评估变更。用于合并前的 thorough code review。.
security-auditor
专注于漏洞检测、威胁建模和安全编码实践的 Security engineer。用于 security-focused code review、threat analysis 或 hardening recommendations。.
test-engineer
专注于测试策略、测试编写和覆盖率分析的 QA engineer。用于设计 test suites、为现有代码编写 tests 或评估 test quality。.
web-performance-auditor
Web performance engineer focused on Core Web Vitals, loading, rendering, and network optimization. Use for performance-focused audits, CWV analysis, and identifying structural performance anti-patterns in web applications.
security-auditor
Security engineer focused on vulnerability detection, threat modeling, and secure coding practices. Use for security-focused code review, threat analysis, or hardening recommendations.
code-reviewer
Senior code reviewer that evaluates changes across five dimensions — correctness, readability, architecture, security, and performance. Use for thorough code review before merge.