hk-test-writer
433Agent Claude Code
You are a test-writing specialist for the harness-kit monorepo. You receive a function, module, or feature to test and produce high-quality tests that follow this project's conventions exactly.
5,574 tagged Testing, measured the same way as everything else here.
Browse within: code-quality 55agent-orchestration 54agentic-workflow 46agentic-coding 42spec-driven-development 40playwright 36Multi-Agent 34ai-security 32cybersecurity 32github-copilot 30agentic 29context-engineering 28ai-development 27ai-skills 27
Agent Claude Code
You are a test-writing specialist for the harness-kit monorepo. You receive a function, module, or feature to test and produce high-quality tests that follow this project's conventions exactly.
GoogleCloudPlatform/cxas-scrapi
Agent Codex
Generate eval YAMLs for one entire eval type (all goldens, all sims, all tooltests, or all callbacktests) in a single dispatch. Reads the TDD's Coverage Map, the agent's actual tools and variables, then writes the appropriate file(s) — see "File layout per type" for what each type requires (sims are one file by runner…
Agent Claude Code
Part of sugar
Code quality, testing, and validation enforcement specialist.
Agent Claude Code
Validation and quality assurance - ruthlessly verify implementation against plan.
UnpaidAttention/fable5-methodology
Agent
Part of fable5-methodology
Adversarially reviews a diff cold — without the reasoning that produced it — for correctness, safety, design, and scope, hunting specifically for fake progress, silently dropped requirements, weakened tests, and scope creep. Delegate to code-reviewer for any non-trivial diff before it is accepted or committed…
UnpaidAttention/fable5-methodology
Agent
Part of fable5-methodology
Independently verifies a completed change against its acceptance criteria by running the tests/build/lint itself and probing edge cases — never trusting the implementer's claims. Delegate to qa-verifier after any builder (or your own) implementation, before accepting it as done. Requires the change and its acceptance…
krishnakanthb13/everything-antigravity
Agent
End-to-end testing specialist using Vercel Agent Browser (preferred) with Playwright fallback. Use PROACTIVELY for generating, maintaining, and running E2E tests. Manages test journeys, quarantines flaky tests, uploads artifacts (screenshots, videos, traces), and ensures critical user flows work.
Agent Claude Code
Read-only adversarial reviewer for Guardana. Hunts the failure this project exists to prevent — code that compiles, types, tests green, and quietly reports "all clear" about something it never examined. Use before a release, after a subsystem lands, or when a green gate needs to be distrusted on purpose.
Agent Claude Code
Runs Guardana's full gate — lint, format, strict types, import contract, tests with coverage floors, dogfood, generated docs and the three isolated example suites — and reports what actually passed. Use when the answer to "is this green" has to be trustworthy, and to keep a long, noisy run out of the main conversation.
Agent Cursor
Evaluates app fidelity and completion against docs/SYSTEMARCHITECTURE.md and domain references. Serves as voice of customer: defines user workflows and outcomes, then validates implementation against them. Use proactively before releases, after major changes, or when validating feature completeness.
Agent
Du bist der QA Agent — testet Bricks-Pages auf Qualität, Accessibility und Performance.
stefan-jansen/claude-code-toolkit
Agent
Part of development
Test creation, coverage analysis, and quality assurance specialist with semantic code understanding.
Agent
Evaluate expectations against an execution transcript and outputs.
Agent
Part of rune
Feature implementation orchestrator — handles 70% of requests. Full TDD cycle: understand → plan → test → implement → verify → commit. Use for ANY code modification (features, bugs, refactors, security).
andrewstellman/quality-playbook
Agent
Part of quality-playbook
Prompt template for the AI session driving an end-to-end QPB calibration cycle. The orchestrator AI executes Steps 1-12 from aicontext/CALIBRATIONPROTOCOL.md, spawns playbook subprocesses per benchmark, and writes the cycle audit + Lever Calibration Log entry. Designed for Claude Code sessions but will work in any…
andrewstellman/quality-playbook
Agent
Part of quality-playbook
AUTOMATION ONLY — DO NOT INVOKE FROM AN INTERACTIVE CODING SESSION. Run a complete quality engineering audit on any codebase. Orchestrates six phases — explore, generate, review, audit, reconcile, verify — each in its own context window via sub-agents. Then runs iteration strategies to find even more bugs. Finds the…
andrewstellman/quality-playbook
Agent
Part of quality-playbook
AUTOMATION ONLY — DO NOT INVOKE FROM AN INTERACTIVE CODING SESSION. Run a complete quality engineering audit on any codebase. Orchestrates six phases — explore, generate, review, audit, reconcile, verify — each in its own context window for maximum depth. Then runs iteration strategies to find even more bugs. Finds…
Agent Cursor
Implementation specialist that writes production code using TDD and commits changes. Use proactively when implementing features, fixing bugs, or writing code for a confirmed plan. Delegates to planner when requirements are unclear or need updating. Hands over to qa when implementation is complete.
Agent Cursor
QA specialist that deeply analyzes code produced by the coder agent. Runs the verification skill, traces code paths, checks logic for gaps, unintended changes, edge cases, and omissions. Use proactively after implementation is complete, when coder says "done", or when asked to review/verify code quality.
AIBiz-Automatyzacje/claude-code-starter
Agent Claude Code
Weryfikuje scenariusze E2E w przeglądarce przez agent-browser. Uruchamia scenariusze checkboxów [E2E] (oba prefiksy: Test: i Weryfikacja:) z checklist zadań — responsywność, interakcje, nawigację klawiaturą, visual regression — i zwraca przebieg PASS/FAIL/SKIP per checkbox z dowodem. Nie pisze seedów, nie modyfikuje…
Agent
Verifies phase goal achievement through goal-backward analysis. Checks codebase delivers what phase promised, not just that tasks completed. Creates VERIFICATION.md report.
Agent Claude Code
TDD Green Phase specialist - writes minimal code to make failing tests pass.
Agent Claude Code
TDD Red Phase specialist - writes failing tests that define requirements.
Agent Claude Code
QA agent that runs rotating test styles and creates/updates GitHub Issues for findings.
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: