sd-test-engineer

sd-test-engineer is an agent for coding agents from Tibsfox/gsd-skill-creator. It costs 0 tokens per session (80 once invoked), scanned A, original, no licence file.

A coding agent focused on automated tests: it writes unit and integration tests, keeps them fast and repeatable, finds flaky tests, and requires regression tests for bug fixes.

In plain words
What is it for?
Use it to add tests, investigate unreliable test results, improve test speed, and check that bug fixes include a failing-then-passing test.
Why use it?
It reduces the risk of broken code reaching users and helps ensure that each bug fix stays fixed.

Agent

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add agents/tibsfox/gsd-skill-creator/sd-test-engineer
Clone the repo
git clone --depth 1 https://github.com/Tibsfox/gsd-skill-creator

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for sd-test-engineer

README.md
[![agentmods](https://agentmods.dev/badge/agents/tibsfox/gsd-skill-creator/sd-test-engineer.svg)](https://agentmods.dev/agents/tibsfox/gsd-skill-creator/sd-test-engineer)
Your own site
<a href="https://agentmods.dev/agents/tibsfox/gsd-skill-creator/sd-test-engineer"><img src="https://agentmods.dev/badge/agents/tibsfox/gsd-skill-creator/sd-test-engineer.svg" alt="Measured on agentmods" height="20"></a>
Per session 0 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 80 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin unknown No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00000 $0.00080
Opus 5 $0.00000 $0.00040
Sonnet 5 $0.00000 $0.00016
Haiku 4.5 $0.00000 $0.00008

Measured 5d ago against content hash a52d69155606, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

sd-test-engineer scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 5d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

examples/cartridges/software-development/agents/sd-test-engineer.md · 7 lines

The source is not reproduced here

A licence we could not identify

The repository carries a LICENSE file, but it is custom or dual enough that GitHub cannot name it and neither can this catalogue. Unknown terms are not permission, so the body is not copied here. Read the licence at the source and decide for yourself.

Read it on GitHub

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 5d ago First seen · 7 lines · 0 tokens per session scan A a52d69155606

Subscribe to this mod's changes

sd-test-engineer is an agent published in the GitHub repository Tibsfox/gsd-skill-creator (69 stars, last pushed 1mo ago), with no licence file. It costs nothing until one of its globs matches a file; then it loads 80 tokens. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other agents, from other repositories

test-writer

Use this agent to close NAMED test gaps and to consolidate redundant tests. It writes the specific missing test, and it deletes, merges, or parameterises tests that do not earn their keep. Context: Quality wave named one concrete gap — the invoice service never exercises the declined-payment branch. user: "The invoice…

Kanevry/session-orchestrator · 281 tokens

orchestrator

Multi-phase development agent. Research > Plan > Implement with validation gates. Use PROACTIVELY when building features that touch >5 files or require architecture decisions.

rohitg00/pro-workflow · 36 tokens

planner

Break down complex tasks into implementation plans before writing code. Use when task touches >5 files, requires architecture decisions, or has unclear requirements.

rohitg00/pro-workflow · 30 tokens

data-engineer

Data specialist — SQL correctness, query performance, schema design, dbt models, pipeline reliability (Airflow / Dagster / Prefect / Fivetran). Use for query review, index planning, migration review, dbt model patterns, and pipeline idempotency.

atuljha23/holocron · 59 tokens

threat-modeler

Structured threat modeling over a diff, a component, or a proposed feature. Walks STRIDE (or LINDDUN for privacy-heavy work) and produces an actionable threat list with mitigations. Use when the task is "what could go wrong here" at the design level, not "is this line safe" (that's @security-reviewer).

atuljha23/holocron · 76 tokens

code-reviewer

Staff-level PR review of the current diff or a specified ref. Use when the user asks for "a review", "a second opinion", or runs /holocron:review. Reviews for correctness, clarity, test quality, and blast radius.

atuljha23/holocron · 54 tokens