Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add skills/rsmdt/the-startup/implement-factorynpx skills add rsmdt/the-startup --skill implement-factorygit clone --depth 1 https://github.com/rsmdt/the-startupWhat it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5 | $0.00037 | $0.02837 |
| Opus 5 | $0.00018 | $0.01418 |
| Sonnet 5 | $0.00007 | $0.00567 |
| Haiku 4.5 | $0.00004 | $0.00284 |
Grade A, and why
implement-factory scanned grade A with 1 finding against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured yesterday.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Makes network callslowCapability
Not a fault in itself. Listed so you know the mod talks to something, and to what.
curl -sf http://localhost:{servicePort}/health && break How it starts
The opening of the file, as written. The whole thing — 323 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Persona
Act as a factory loop orchestrator that implements specifications by spawning isolated subagents. You control information flow between code agents and evaluation agents. You never implement code directly.
Implementation Target: $ARGUMENTS
Interface
Unit { id: string // e.g., "ve1" title: string dependencies: string[] // unit IDs this unit depends on status: pending | in_progress | completed | failed iteration: number // current retry count (starts at 0) failureSummaries: string[] // one-line summaries from last evaluation }
ExecutionGroup { number: number mode: parallel | sequential unitIds: string[] }
EvaluationResult { unitId: string satisfaction: number // 0.0 - 1.0 passed: string[] // scenario names that passed failed: FailedScenario[] }
FailedScenario { name: string summary: string // one-line observable symptom failCount: string // e.g., "3/3 failures" }
Manifest { title: string status: pending | in_progress | completed | failed threshold: number // e.g., 0.90 maxIterations: number // e.g., 5 units: Unit[] executionGroups: ExecutionGroup[] }
State { target = $ARGUMENTS specDirectory: string // resolved .start/specs/NNN-name/ path manifest: Manifest servicePort: number // discovered from project instructions or package.json startCommand: string // discovered from project instructions or package.json serviceProcess: active | stopped }
Constraints
Always:
- Delegate ALL implementation to code agents and ALL evaluation to evaluation agents — spawn each as an isolated specialist subagent.
- Construct each agent's prompt using the templates in reference/code-agent.md and reference/eval-agent.md.
- Enforce information barriers: code agents never see scenarios; evaluation agents never see source code or unit specs.
- Filter failure feedback to one-line summaries only — never pass scenario text or full evaluation output to code agents.
- Start the service once per execution group; keep it running across all evaluations in that group.
- Health-check before every evaluation phase.
- Restart the service only if a code agent changed server-side code on retry.
- Update manifest.md checkboxes and frontmatter status as units complete.
- Skip already-completed units when resuming an interrupted manifest.
- Present satisfaction metrics to the user after each evaluation.
- Escalate to the user when max iterations is reached for any unit.
- Use the validate skill in constitution mode at group boundaries if a CONSTITUTION.md exists at the project root.
What ships with it
40 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.
- evals/evals.json 7.1 KB
- evals/fixtures/eval-1-happy-path/manifest.md 338 B
- evals/fixtures/eval-1-happy-path/scenarios/dm1/health-status-fields.md 346 B
- evals/fixtures/eval-1-happy-path/scenarios/rl1/rate-limit-exceeded.md 294 B
- evals/fixtures/eval-1-happy-path/scenarios/ve1/endpoint-returns-200.md 291 B
- evals/fixtures/eval-1-happy-path/scenarios/ve1/invalid-method-405.md 186 B
- evals/fixtures/eval-1-happy-path/units/dm1.md 488 B
- evals/fixtures/eval-1-happy-path/units/rl1.md 421 B
- evals/fixtures/eval-1-happy-path/units/ve1.md 496 B
- evals/fixtures/eval-2-retry/manifest.md 216 B
- evals/fixtures/eval-2-retry/scenarios/ve1/age-out-of-range.md 315 B
- evals/fixtures/eval-2-retry/scenarios/ve1/invalid-email.md 309 B
- evals/fixtures/eval-2-retry/scenarios/ve1/valid-input.md 249 B
- evals/fixtures/eval-2-retry/units/ve1.md 559 B
- evals/fixtures/eval-3-max-iterations/manifest.md 204 B
- evals/fixtures/eval-3-max-iterations/scenarios/ve1/returns-version.md 246 B
- evals/fixtures/eval-3-max-iterations/units/ve1.md 282 B
- evals/fixtures/eval-4-parallel/manifest.md 248 B
- evals/fixtures/eval-4-parallel/scenarios/ep1/ping-returns-pong.md 187 B
- evals/fixtures/eval-4-parallel/scenarios/ep2/echo-missing-body.md 164 B
- evals/fixtures/eval-4-parallel/scenarios/ep2/echo-returns-body.md 226 B
- evals/fixtures/eval-4-parallel/units/ep1.md 247 B
- evals/fixtures/eval-4-parallel/units/ep2.md 280 B
- evals/fixtures/eval-5-resume/manifest.md 273 B
- evals/fixtures/eval-5-resume/scenarios/dm1/health-data.md 170 B
- evals/fixtures/eval-5-resume/scenarios/ve1/health-endpoint.md 209 B
- evals/fixtures/eval-5-resume/units/dm1.md 275 B
- evals/fixtures/eval-5-resume/units/ve1.md 303 B
- evals/fixtures/eval-6-e2e-stubs/manifest.md 251 B
- evals/fixtures/eval-6-e2e-stubs/scenarios/dm1/e2e-stubs.md 430 B
- evals/fixtures/eval-6-e2e-stubs/scenarios/dm1/echo-response.md 195 B
- evals/fixtures/eval-6-e2e-stubs/scenarios/ve1/e2e-stubs.md 1.0 KB
- evals/fixtures/eval-6-e2e-stubs/scenarios/ve1/echo-missing-message.md 213 B
- evals/fixtures/eval-6-e2e-stubs/scenarios/ve1/echo-success.md 278 B
- evals/fixtures/eval-6-e2e-stubs/units/dm1.md 372 B
- evals/fixtures/eval-6-e2e-stubs/units/ve1.md 347 B
- examples/output-example.md 2.8 KB
- reference/code-agent.md 3.0 KB
- reference/eval-agent.md 3.6 KB
- reference/output-format.md 2.1 KB
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- yesterday First seen · 323 lines · 37 tokens per session scan A ad6840df7710
implement-factory is a skill published in the GitHub repository rsmdt/the-startup (511 stars, last pushed 28d ago), licensed MIT. It adds 37 tokens to every session and 2,837 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 1 finding (makes network calls). No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.
Other skills, from other repositories
postgres-database-migration
Use this skill for planning, testing, and safely executing PostgreSQL schema migrations — especially when working with production data or shared databases. Trigger when user asks to: Test a schema migration before applying it to production Add, remove, or rename columns safely on a live table Change a column's data…
design-postgis-tables
Comprehensive PostGIS spatial table design reference covering geometry types, coordinate systems, spatial indexing, and performance patterns for location-based applications.
setup-timescaledb-hypertables
Use this skill when creating database schemas or tables for Timescale, TimescaleDB, TigerData, or Tiger Cloud, especially for time-series, IoT, metrics, events, or log data. Use this to improve the performance of any insert-heavy table. Trigger when user asks to: Create or design SQL schemas/tables AND…
migrate-postgres-tables-to-hypertables
Use this skill to migrate identified PostgreSQL tables to Timescale/TimescaleDB hypertables with optimal configuration and validation. Trigger when user asks to: Migrate or convert PostgreSQL tables to hypertables Execute hypertable migration with minimal downtime Plan blue-green migration for large tables Validate…
pgvector-semantic-search
Use this skill for setting up vector similarity search with pgvector for AI/ML embeddings, RAG applications, or semantic search. Trigger when user asks to: Store or search vector embeddings in PostgreSQL Set up semantic search, similarity search, or nearest neighbor search Create HNSW or IVFFlat indexes for vectors…
postgres-hybrid-text-search
Use this skill to implement hybrid search combining BM25 keyword search with semantic vector search using Reciprocal Rank Fusion (RRF). Trigger when user asks to: Combine keyword and semantic search Implement hybrid search or multi-modal retrieval Use BM25/pgtextsearch with pgvector together Implement RRF (Reciprocal…