test-team-leader

test-team-leader is an agent for coding agents from nrslib/takt. It costs 0 tokens per session (66 once invoked), scanned A, original, MIT.

An agent role that leads end-to-end testing, meaning tests of a complete application flow from start to finish.

In plain words
What is it for?
Use it to split an end-to-end testing task into two self-contained, executable subtasks.
Why use it?
It turns a broad testing task into two separate pieces that can be worked on independently.

Agent

About the project

TAKT is a command-line workflow engine for coordinating AI coding agents through defined planning, implementation, review, and repair steps. Development teams describe agent roles, permissions, checkpoints, and outputs in YAML, while TAKT runs tasks in isolated worktrees and records their results; the catalogue entries provide agents, instructions, and skills for these workflows.

nrslib/takt · 1,328 stars · on GitHub

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add agents/nrslib/takt/test-team-leader
Clone the repo
git clone --depth 1 https://github.com/nrslib/takt

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for test-team-leader

README.md
[![agentmods](https://agentmods.dev/badge/agents/nrslib/takt/test-team-leader.svg)](https://agentmods.dev/agents/nrslib/takt/test-team-leader)
Your own site
<a href="https://agentmods.dev/agents/nrslib/takt/test-team-leader"><img src="https://agentmods.dev/badge/agents/nrslib/takt/test-team-leader.svg" alt="Measured on agentmods" height="20"></a>
Per session 0 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 66 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00000 $0.00066
Opus 5 $0.00000 $0.00033
Sonnet 5 $0.00000 $0.00013
Haiku 4.5 $0.00000 $0.00007

Measured 5d ago against content hash aa9759a274c7, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

test-team-leader scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 5d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

e2e/fixtures/agents/test-team-leader.md · 10 lines

What it actually says

E2E Test Team Leader

You are a team leader for E2E testing. Your job is to decompose a task into independent subtasks.

Instructions

  • Analyze the task and split it into 2 independent parts
  • Each part should be self-contained and executable independently
  • Keep instructions concise and actionable
Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 5d ago First seen · 10 lines · 0 tokens per session scan A aa9759a274c7

Subscribe to this mod's changes

test-team-leader is an agent published in the GitHub repository nrslib/takt (1,328 stars, last pushed 3d ago), licensed MIT. It costs nothing until one of its globs matches a file; then it loads 66 tokens. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other agents, from other repositories

codex-qa-tester

Manual QA profile for browser testing, workflow verification, and regression checks.

Waishnav/devspace · 21 tokens

ux-evaluator

Use this agent for read-only UX evaluation of test-runner driver artifacts (Playwright AX-tree snapshots, screenshots, console output). Applies the 4-check UX rubric (onboarding-step-count ≤7, axe-violations critical/serious, console-errors visible to user, Apple-Liquid-Glass .glassEffect() conformance on SwiftUI 26+)…

Kanevry/session-orchestrator · 199 tokens

rondoflow-reviewer

Reviews a RondoFlow code change against this project's specific conventions and security rules (childprocess spawn safety, the { success, error } API envelope, per-user ownership/IDOR, Zod boundary validation, immutability, i18n parity, file/function size limits, Claude-auth handling). Use after writing or before…

rondoflow/rondoflow · 112 tokens

rondoflow-i18n-translator

Translates RondoFlow UI strings into Slovak (sk) and Spanish (es) and writes them into the locale catalogs, keeping key/interpolation parity with English so the catalog test passes. Use when English keys were added or changed and the sk/es catalogs need to be filled in, or when asked to translate/localize UI strings.…

rondoflow/rondoflow · 91 tokens

playwright-test-healer

Use this agent when you need to debug and fix failing Playwright tests.

saltbo/agent-kanban · 20 tokens

health-sweeper

Samples teamctl's running state — mailbox.db size, live tmux session count vs configured roster, open painpoint counts, recent supervisor log lines — and flags anomalies against the known baseline. Use when Otto wants a passive background health sweep of the dogfood team. Returns a 2-4 line summary (all-nominal or the…

Alireza29675/teamctl · 92 tokens