a5c-ai/babysitter

Babysitter enforces obedience on agentic workforces and enables them to manage extremely complex tasks and workflows through deterministic, hallucination-free self-orchestration

About the project

Babysitter is a workflow engine for AI coding agents that enforces predefined steps, quality checks, human approvals, and decision records. It is used to coordinate complex, repeatable agent workflows across supported coding tools. The catalogue contains skills, agents, instructions, settings, a plugin, and an MCP integration for its workflow.

This repository also configures its own agents. See what babysitter tells them →

1.8kStars on the repository
186Mods indexed here, across every type
7d agoLast push, which is what freshness is scored on
MITLicence, which decides whether bodies are shown

a5c-ai/babysitter

Skill Claude Code

State capture and restore across context window compactions. Monitors usage thresholds and serializes quality, task, and spec state for seamless continuation.

not rated 1.8k +16 7d ago A SkillSpector: pass 32 tokens original MIT

persistent-memory

98

a5c-ai/babysitter

Skill Claude Code

Observation capture and retrieval across sessions. Stores decisions, discoveries, and bugfix patterns. Searchable via tags and relevance scoring.

not rated 1.8k +16 7d ago A SkillSpector: pass 28 tokens original MIT

quality-hooks

99

a5c-ai/babysitter

Skill Claude Code

Language-specific auto-lint/format/typecheck pipeline. Supports Python (ruff+pyright), TypeScript (prettier+eslint+tsc), Go (gofmt+golangci-lint). Auto-fix and convergence loops.

not rated 1.8k +16 7d ago A SkillSpector: warn 53 tokens original MIT

a5c-ai/babysitter

Skill Claude Code

Specification creation and management for the Pilot Shell methodology. Covers semantic search, clarifying questions, structured spec generation, and iterative refinement.

not rated 1.8k +16 7d ago A SkillSpector: pass 30 tokens original MIT

strict-tdd

101

a5c-ai/babysitter

Skill Claude Code

Strict RED->GREEN->REFACTOR test-driven development with enforcement. Never write production code before a failing test. Atomic commits per TDD cycle.

not rated 1.8k +16 7d ago A SkillSpector: pass 34 tokens original MIT

a5c-ai/babysitter

Skill Claude CodeCodex

Verify all phases are complete with weighted quality scoring before allowing session exit.

not rated 1.8k +16 7d ago A SkillSpector: pass 0 tokens original MIT

error-logging

103

a5c-ai/babysitter

Skill Claude CodeCodex

Log all errors with full context, detect patterns, and suggest approach mutations to avoid repeated failures.

not rated 1.8k +16 7d ago A SkillSpector: pass 0 tokens original MIT

findings-capture

104

a5c-ai/babysitter

Skill Claude CodeCodex

Capture and persist research findings, discoveries, and decision rationale to findings.md.

not rated 1.8k +16 7d ago A SkillSpector: pass 0 tokens original MIT

plan-creation

105

a5c-ai/babysitter

Skill Claude CodeCodex

Create a structured taskplan.md with phases, goals, and checkbox tracking for persistent planning.

not rated 1.8k +16 7d ago A SkillSpector: pass 0 tokens original MIT

session-recovery

106

a5c-ai/babysitter

Skill Claude CodeCodex

Detect and recover previous planning sessions, reconstructing lost context from persistent planning files.

not rated 1.8k +16 7d ago A SkillSpector: pass 0 tokens original MIT

brainstorming

107

a5c-ai/babysitter

Skill Claude Code

Clarify vague requirements through exploratory questioning and option generation before committing to research or implementation.

not rated 1.8k +16 7d ago A SkillSpector: pass 21 tokens original MIT

codebase-research

108

a5c-ai/babysitter

Skill Claude Code

Systematic codebase exploration following the Iron Law - understand the problem before exploring code. Four phases with file-finder and web-researcher agents.

not rated 1.8k +16 7d ago A SkillSpector: pass 35 tokens original MIT

a5c-ai/babysitter

Skill Claude Code

Create Architecture Decision Records (ADRs) documenting significant technical choices with context, options, consequences, and sequential numbering.

not rated 1.8k +16 7d ago A SkillSpector: pass 27 tokens original MIT

finishing-work

110

a5c-ai/babysitter

Skill Claude Code

Final completion discipline including summary generation, plan document updates, and confirmation that all success criteria from the original plan are satisfied.

not rated 1.8k +16 7d ago A SkillSpector: pass 28 tokens original MIT

plan-implementation

111

a5c-ai/babysitter

Skill Claude Code

Disciplined execution of approved plans with step-by-step verification, phase checkpoints, failure investigation, and mandatory code/security reviews.

not rated 1.8k +16 7d ago A SkillSpector: pass 29 tokens original MIT

plan-writing

112

a5c-ai/babysitter

Skill Claude Code

Transform research findings into actionable implementation plans with stakes-based rigor, test-first strategy, and granular task decomposition.

not rated 1.8k +16 7d ago A SkillSpector: pass 24 tokens original MIT

security-review

113

a5c-ai/babysitter

Skill Claude Code

Security vulnerability assessment identifying OWASP risks, injection vectors, authentication issues, and data exposure with severity classification.

not rated 1.8k +16 7d ago A SkillSpector: pass 24 tokens original MIT

systematic-debugging

114

a5c-ai/babysitter

Skill Claude Code

Structured debugging methodology using hypothesis-driven investigation, log analysis, and bisection to isolate and resolve defects.

not rated 1.8k +16 7d ago A SkillSpector: pass 26 tokens original MIT

verification

115

a5c-ai/babysitter

Skill Claude Code

Verification-before-completion discipline ensuring all success criteria are met, tests pass, and reviews complete before declaring work done.

not rated 1.8k +16 7d ago A SkillSpector: pass 25 tokens original MIT

agent-booster

116

a5c-ai/babysitter

Skill Claude Code

WASM-based instant code transforms for simple tasks, achieving 352x speedup over LLM inference with zero cost.

not rated 1.8k +16 7d ago A SkillSpector: pass 28 tokens original MIT

anti-drift

117

a5c-ai/babysitter

Skill Claude Code

Hierarchical coordination and drift detection with frequent checkpoints, shared memory coherence validation, role specialization enforcement, and short task cycles.

not rated 1.8k +16 7d ago A SkillSpector: pass 28 tokens original MIT

consensus-mechanisms

118

a5c-ai/babysitter

Skill Claude Code

Multi-protocol consensus for agent swarms supporting Raft leader election, Byzantine fault tolerance, Gossip state propagation, and CRDT conflict-free merging.

not rated 1.8k +16 7d ago A SkillSpector: pass 35 tokens original MIT

security-hardening

119

a5c-ai/babysitter

Skill Claude Code

AIDefence security layer with prompt injection blocking, input validation, sandboxed execution, output sanitization, and STRIDE threat modeling.

not rated 1.8k +16 7d ago A SkillSpector: pass 32 tokens original MIT

self-optimization

120

a5c-ai/babysitter

Skill Claude Code

SONA self-optimizing neural architecture with ReasoningBank trajectory learning, EWC++ anti-forgetting, and reinforcement learning feedback loops.

not rated 1.8k +16 7d ago A SkillSpector: pass 32 tokens original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: