Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx skills add cisco-open/ai-harness-toolkit --skill ai-harness-setupgit clone --depth 1 https://github.com/cisco-open/ai-harness-toolkitWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/cisco-open/ai-harness-toolkit/ai-harness-setup)<a href="https://agentmods.dev/skills/cisco-open/ai-harness-toolkit/ai-harness-setup"><img src="https://agentmods.dev/badge/skills/cisco-open/ai-harness-toolkit/ai-harness-setup/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/skills/cisco-open/ai-harness-toolkit/ai-harness-setup"><img src="https://agentmods.dev/badge/skills/cisco-open/ai-harness-toolkit/ai-harness-setup.svg" alt="Reviewed on agentmods" width="80" height="20"></a>- NVIDIA SkillSpector warn
SkillSpector: 1 finding, up to medium
These are SkillSpector’s own severities. On a checked sample its high-severity flags on skills were ~96% false positives — a documented command, a public API, a “never do X” rule — so we show them as a caution to read, not a verdict. Why →
- medium MCP Rug Pull · line 184 npx commands without a version suffix (e.g. @1.0.0) create a rug-pull risk if the upstream server is compromised and publishes a malicious update.Fix: Pin the version: npx @scope/[email protected]
What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00123 | $0.04945 |
| Opus 5 | $0.00062 | $0.02472 |
| Sonnet 5 | $0.00025 | $0.00989 |
| Haiku 4.5 | $0.00012 | $0.00494 |
Grade A, and why
ai-harness-setup scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 9d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 284 lines — stays where its author put it; the contents beside it link to each section on GitHub.
AI Harness Setup
Use this skill to bootstrap a repository toward a durable agent-driven workflow: deep inspection first, then APM as the dependency and skill delivery layer, spec-driven development for change management (OpenSpec by default, or the repo's existing system), deterministic checks matched to the detected stack, AI skills for review and delivery, a docs tree that explains the whole system, and AI IDE configuration for every IDE the team uses.
Quick Start
- Inspect first. Search the repo's manifests, lockfiles, CI configs, framework configs, and existing docs to classify the full stack before changing anything. Record: languages, package managers, type checkers, test runners, linters, CI system, monorepo layout, and existing tooling.
- Read the matching stack guide based on detection results:
references/stacks/javascript-typescript/stack.mdreferences/stacks/python/stack.mdreferences/stacks/java/stack.md
- If the repo uses a major framework, also read the matching framework guide:
references/stacks/javascript-typescript/framework-react.mdreferences/stacks/javascript-typescript/framework-angular.mdreferences/stacks/java/framework-spring-boot.md
- Ask whether the team wants GitHub Agentic Workflows enabled in this repository.
- Ask whether deterministic checks should be
enforcedoradvisorybefore configuring any checks. Carry that answer through all downstream local scripts, CI wiring, and git hooks. - If GitHub Agentic Workflows are enabled, read
references/github-agentic-workflows.mdbefore installing AI workflow tooling. - Read
references/ai-tooling.mdbefore initializing APM or installing package-backed skills. - Read
references/mcp-servers.mdbefore adding extra MCP servers toapm.yml. - Read
references/openspec.mdfor change management setup. - Read
references/deterministic-checks-core.mdand the matching per-stackdeterministic-scans.mdtogether -- they form one phase. - Read
references/dependabot.mdwhen adding Dependabot configuration. - Read
references/docs-bootstrap.mdwhen creating the docs section that explains the workflow. - Read
references/opencode.mdwhen OpenCode orocxsupport is requested. - Verify is mandatory. After all setup steps finish, run the Verify step in Workflow Step 10 to audit the changeset against the harness instructions, fix every gap found, and produce a verify report. Never skip this step.
What ships with it
24 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.
- README.md 4.2 KB
- references/ai-tooling.md 3.6 KB
- references/ci-wiring.md 7.2 KB
- references/dependabot.md 2.5 KB
- references/deterministic-checks-core.md 5.4 KB
- references/deterministic-scans.md 1.4 KB
- references/docs-bootstrap.md 13 KB
- references/github-agentic-workflows.md 4.0 KB
- references/mcp-servers.md 2.8 KB
- references/opencode.md 4.0 KB
- references/openspec.md 3.7 KB
- references/stacks/java/deterministic-scans.md 4.2 KB
- references/stacks/java/framework-spring-boot.md 1.7 KB
- references/stacks/java/stack.md 2.2 KB
- references/stacks/javascript-typescript/deterministic-scans.md 4.6 KB
- references/stacks/javascript-typescript/framework-angular.md 1.8 KB
- references/stacks/javascript-typescript/framework-react.md 2.3 KB
- references/stacks/javascript-typescript/stack.md 5.3 KB
- references/stacks/python/deterministic-scans.md 4.5 KB
- references/stacks/python/framework-django.md 1.6 KB
- references/stacks/python/framework-fastapi.md 1.4 KB
- references/stacks/python/stack.md 5.0 KB
- references/templates/AGENTS.md 2.5 KB
- references/verify-harness.md 10 KB
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 9d ago First seen · 284 lines · 123 tokens per session scan A 382eca7949ad
ai-harness-setup is a skill published in the GitHub repository cisco-open/ai-harness-toolkit (11 stars, last pushed 13d ago), licensed Apache-2.0. It adds 123 tokens to every session and 4,945 once invoked, about $0.0006 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other skills, from other repositories
systematic-debugging
Use when encountering any bug, test failure, or unexpected behavior, before proposing fixes.
local-ai-agents
Build local-first AI agents that run entirely on a developer workstation with Microsoft Foundry Local and Qwen function-calling models. Covers Small Language Models (SLMs), the OpenAI-compatible local endpoint, sandboxed local tools, local RAG with Chroma, local MCP servers, hybrid cloud/local routing, and the…
next-cache-components-adoption
Turn on Cache Components in a Next.js app and resolve the blocking routes it surfaces. Use when the user wants to enable, adopt, or migrate to Cache Components, flip the cacheComponents flag, work through a flood of blocking-prerender / instant validation errors, run the cache-components-instant-false codemod, or…
insight-error-page
Write or audit an insight-kind error page for the Next.js dev overlay. Use when creating a new errors/ .mdx page, auditing an existing one, or checking that a page matches the framework fix cards. Covers page structure, title alignment, FixCard cards with Copy prompt button, code snippets, terminology verification…
next-cache-components-optimizer
Drive a Next.js route to instant navigation by setting up an agentic loop, under Cache Components / PPR, on initial load (hard navigation) and client-side navigation (soft navigation). Encode the goal as a failing @next/playwright instant() e2e and work it to green, one verified route at a time; the shipped test then…
next-partial-prefetching-adoption
Turn on Partial Prefetching in a Next.js app and work through the insights it surfaces. Use when the user wants to enable or adopt Partial Prefetching, flip the partialPrefetching flag, opt routes in with export const prefetch = 'partial', audit Link prefetch={true} behavior, preserve existing prefetched UI with…