Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add skills/nick-pape/grackle/test-acp-runtimenpx skills add nick-pape/grackle --skill test-acp-runtimegit clone --depth 1 https://github.com/nick-pape/grackleWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/nick-pape/grackle/test-acp-runtime)<a href="https://agentmods.dev/skills/nick-pape/grackle/test-acp-runtime"><img src="https://agentmods.dev/badge/skills/nick-pape/grackle/test-acp-runtime.svg" alt="Measured on agentmods" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00064 | $0.00784 |
| Opus 5 | $0.00032 | $0.00392 |
| Sonnet 5 | $0.00013 | $0.00157 |
| Haiku 4.5 | $0.00006 | $0.00078 |
Grade A, and why
test-acp-runtime scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 6d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 46 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Test the ACP runtimes
How to spawn the ACP-variant runtimes — claude-code-acp, codex-acp, copilot-acp — against an isolated test server. Assumes a server from /launch-grackle (with GRACKLE_URL + GRACKLE_API_KEY exported).
Why ACP matters: the native runtimes (
claude-code/copilot/codex) are slated to be deprecated in favor of these ACP variants. ACP is a standard protocol with a uniformtool_call_updatestatus(completed/failed), so one adapter (runtime-acp) covers all agents — and unlike nativeclaude-code(which emits synthetic emptytool_results and can't surface failures),claude-code-acpreports real tool failures.
Model names
ACP runtimes are listed as model "(agent-selected)" by grackle runtimes, but a persona still requires a model (spawn errors with FailedPrecondition: persona has no model configured otherwise):
| Runtime | Model to pass |
|---|---|
claude-code-acp |
sonnet / opus / haiku (runs Claude underneath) — validated with sonnet |
codex-acp |
a Codex model (same ChatGPT-account gating as native — use gpt-5.5; see /test-codex-runtime) |
copilot-acp |
a Copilot model, e.g. claude-sonnet-4.5 (see /test-copilot-runtime) |
Spawn it
grackle persona create "ACP Tester" --runtime claude-code-acp --model sonnet --prompt "You are a test agent."
grackle spawn local "Run exactly this one shell command and nothing else, then stop: cat /nonexistent_file_xyz" --persona acp-tester
What it proves (confirmed live 2026-05-29)
- Spawns cleanly. The earlier
Invalid permissions.defaultMode: autofailure (#1366) is fixed by #1370 —claude-code-acpnow isolates itsCLAUDE_CONFIG_DIRfrom your personal~/.claude(which may carry the interactive-onlydefaultMode: "auto"). - Surfaces real tool failures. A failing command (
cat /nonexistent) produced atool_resultwithcontent {is_ok:false}+"tool_error":trueand the realExit code 1/No such filetext — via the adapter'sstatus === "failed"→toolErrormapping (packages/runtime-acp/src/acp.ts). This is the case nativeclaude-codecannot express.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 6d ago First seen · 46 lines · 64 tokens per session scan A 737b7fb8cf28
test-acp-runtime is a skill published in the GitHub repository nick-pape/grackle (21 stars, last pushed 2mo ago), licensed MIT. It adds 64 tokens to every session and 784 once invoked, about $0.0003 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.
Other skills, from other repositories
copilot-pr-review-loop
Drive a GitHub pull request through repeated rounds of Copilot code review until convergence. Use when the user asks to "request Copilot review", "run a Copilot review loop", iterate on Copilot feedback, or wants automated triage-and-respond on Copilot PR comments. Covers re-request mechanics, open-thread filtering…
upstream-sync
Periodically sync new commits from microsoft/terminal into this manually-forked intelligent-terminal repo by cherry-picking commit-by-commit onto a dated sync branch, auto-skipping revert pairs and empty commits, auto-resolving known take-upstream files, and stopping cleanly on genuine conflicts. The agent (you…
release-notes
Generate user-facing release notes for Intelligent Terminal. Use when asked to write release notes, changelog, what-is-new summary, or prepare a release. Compares git commits between releases, looks up PR-linked issues and community contributors, then outputs formatted notes with "Verbed + Impact + Scenario" style…
add-acp-agent-support
Add first-class support for an ACP-compatible agent CLI to Intelligent Terminal. Use when integrating a new built-in AI agent, ACP server command, authentication flow, model selection, interactive delegation, session hooks, onboarding, Settings, branding, GPO policy, documentation, tests, build, deployment, or live…
pr-integration-test
Design, implement, and validate Intelligent Terminal integration tests for a target pull request or regression. Use when asked to add PR integration tests, convert a bug fix into E2E coverage, prove existing behavior still works, map tests to the release checklist, or verify E2E reports mark checklist cases complete.
opentag
Deploy or operate OpenTag's supported paired setup when a user needs to bootstrap the self-hosted Docker Compose Control Plane and Slack Source App, configure or pair a local ACP Runner with a GitHub Project Target, start the Runner service, verify readiness, or diagnose a Slack mention that did not complete.