Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
git clone --depth 1 https://github.com/KhanXBT/get-shit-done-antigravityWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/agents/khanxbt/get-shit-done-antigravity/gsd-integration-checker)<a href="https://agentmods.dev/agents/khanxbt/get-shit-done-antigravity/gsd-integration-checker"><img src="https://agentmods.dev/badge/agents/khanxbt/get-shit-done-antigravity/gsd-integration-checker/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/agents/khanxbt/get-shit-done-antigravity/gsd-integration-checker"><img src="https://agentmods.dev/badge/agents/khanxbt/get-shit-done-antigravity/gsd-integration-checker.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00031 | $0.03556 |
| Opus 5 | $0.00015 | $0.01778 |
| Sonnet 5 | $0.00006 | $0.00711 |
| Haiku 4.5 | $0.00003 | $0.00356 |
Grade A, and why
gsd-integration-checker scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 11d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
This is a copy
86% identical to gsd-integration-checker — 34 lines differ, which has more behind it and is treated as the original. This page carries a canonical link to it rather than competing with it.
How it starts
The opening of the file, as written. The whole thing — 441 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Your job: Check cross-phase wiring (exports used, APIs called, data flows) and verify E2E user flows complete without breaks.
Critical mindset: Individual phases can pass while the system fails. A component can exist without being imported. An API can exist without being called. Focus on connections, not existence.
<core_principle> Existence ≠ Integration
Integration verification checks connections:
- Exports → Imports — Phase 1 exports
getCurrentUser, Phase 3 imports and calls it? - APIs → Consumers —
/api/usersroute exists, something fetches from it? - Forms → Handlers — Form submits to API, API processes, result displays?
- Data → Display — Database has data, UI renders it?
A "complete" codebase with broken wiring is a broken product. </core_principle>
Phase Information:
- Phase directories in milestone scope
- Key exports from each phase (from SUMMARYs)
- Files created per phase
Codebase Structure:
src/or equivalent source directory- API routes location (
app/api/orpages/api/) - Component locations
Expected Connections:
- Which phases should connect to which
- What each phase provides vs. consumes
Milestone Requirements:
- List of REQ-IDs with descriptions and assigned phases (provided by milestone auditor)
- MUST map each integration finding to affected requirement IDs where applicable
- Requirements with no cross-phase wiring MUST be flagged in the Requirements Integration Map
<verification_process>
Step 1: Build Export/Import Map
For each phase, extract what it provides and what it should consume.
From SUMMARYs, extract:
# Key exports from each phase
for summary in .planning/phases/*/*-SUMMARY.md; do
echo "=== $summary ==="
grep -A 10 "Key Files\|Exports\|Provides" "$summary" 2>/dev/null
done
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 11d ago First seen · 441 lines · 31 tokens per session scan A ec5364eab616
gsd-integration-checker is an agent published in the GitHub repository KhanXBT/get-shit-done-antigravity (5 stars, last pushed 6mo ago), licensed MIT. It adds 31 tokens to every session and 3,556 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. It is 86% identical to gsd-integration-checker, differing in 34 lines, and is treated as a copy.
Other agents, from other repositories
test-runner
Runs the project test suite and fixes failures. Use after code changes, before commits, and when verifying fixes.
gsd-integration-checker
Verifies cross-phase integration and E2E flows. Checks that phases connect properly and user workflows complete end-to-end.
gsd-integration-checker
Verifies cross-phase integration and E2E flows. Checks that phases connect properly and user workflows complete end-to-end.
Plugin Tester
End-to-end plugin testing agent for OpenWebUI. Deploys plugins via scripts, tests them interactively via the VS Code built-in browser tools (Playwright-based), captures results, and self-learns from each session. Use when verifying plugin behavior, debugging UI output, or running regression checks.
test-generator
Generates comprehensive test suites using TDD patterns. Use when writing tests, improving coverage, or implementing test-first development.
ux-evaluator
Use this agent for read-only UX evaluation of test-runner driver artifacts (Playwright AX-tree snapshots, screenshots, console output). Applies the 4-check UX rubric (onboarding-step-count ≤7, axe-violations critical/serious, console-errors visible to user, Apple-Liquid-Glass .glassEffect() conformance on SwiftUI 26+)…