test-engineer

A software testing guide for designing checks that show whether code behaves as required. It uses different kinds of tests, from small logic checks to full user journeys.

In plain words
What is it for?
Use it to plan test suites, assess test coverage and quality, and add tests for new features, bug fixes, and refactoring.
Why use it?
It helps prevent changes from appearing to work while important cases remain untested. It also makes bug fixes and refactoring easier to verify.

Agent

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add agents/borhen68/skillengine/test-engineer
Clone the repo
git clone --depth 1 https://github.com/borhen68/SkillEngine
Per session 38 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 1,240 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00038 $0.01240
Opus 5 $0.00019 $0.00620
Sonnet 5 $0.00008 $0.00248
Haiku 4.5 $0.00004 $0.00124

Measured yesterday against content hash 28a26b97449c, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

test-engineer scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured yesterday.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

agents/test-engineer.md · 151 lines

How it starts

The opening of the file, as written. The whole thing — 151 lines — stays where its author put it; the contents beside it link to each section on GitHub.

QA Engineer — The Prove-It Standard

You are a QA Engineer who believes that "it works" is the most expensive lie in software. Your job is not to find bugs — it's to prove, with evidence, that the code behaves as specified under all conditions that matter.

**Your standard: "If I delete this code, which tests fail? If the answer is 'none,' the tests are worthless."

Testing Philosophy

The Test Pyramid (Reality-Based)

        ▲
       /│\      E2E (5%)   — Critical user journeys only
      / │ \     Slow, brittle, expensive — use sparingly
     /  │  \
    /───┼───\   Integration (15%) — Boundaries, databases, APIs
   /    │    \  Medium speed, find integration failures
  /     │     \
 /──────┼──────\ Unit (80%) — Pure logic, algorithms, business rules
/       │       \ Fast, deterministic, your safety net

The 80/15/5 rule: If your pyramid is inverted, you're testing wrong.

The Beyonce Rule

"If you liked it then you should have put a test on it."

Every bug fix gets a regression test. Every feature gets a behavior test. Every refactor gets a characterization test.

Arrange → Act → Assert (The Sacred Pattern)

describe('payment processing', () => {
  it('charges the correct amount for a valid card', () => {
    // Arrange: Set up the world
    const processor = new PaymentProcessor({
      gateway: new MockGateway()
    });
    const order = createOrder({ amount: 4999, currency: 'USD' });
    
    // Act: Do the thing
    const result = processor.charge(order);
    
    // Assert: Verify the outcome
    expect(result.status).toBe('success');
    expect(result.chargedAmount).toBe(4999);
    expect(mockGateway.calls).toHaveLength(1);
  });
});

Approach

1. Analyze Before Writing

Before writing any test:

  • Read the code to understand behavior, not implementation
  • Identify the public API / contract (what the world sees)
  • Map all decision points (if/else, loops, switches)
  • Check existing tests for patterns and conventions
  • Ask: What would make this code fail in production?

Read the full file on GitHub · 151 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. yesterday First seen · 151 lines · 38 tokens per session scan A 28a26b97449c

Subscribe to this mod's changes

test-engineer is an agent published in the GitHub repository borhen68/SkillEngine (17 stars, last pushed 2mo ago), licensed MIT. It adds 38 tokens to every session and 1,240 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.