Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add instructions/srnichols/plan-forge/testinggit clone --depth 1 https://github.com/srnichols/plan-forgeWhat it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5 | $0.03216 | $0.03216 |
| Opus 5 | $0.01608 | $0.01608 |
| Sonnet 5 | $0.00643 | $0.00643 |
| Haiku 4.5 | $0.00322 | $0.00322 |
Grade A, and why
plan-forge testing.instructions.md scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured yesterday.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 230 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Testing Instructions
When this loads: every time you edit, write, debug, or report on a test file. Sister script:
scripts/audit/test-smells.mjs— mechanical scan for the patterns this file rules against.
The 6 Rules
1. Time-sensitive tests must declare tolerance OR use fake timers
The single biggest source of CI flake. Phase 41 Slice 5 reference incident: timeline-core cache-invalidation test used a +5ms tolerance — too tight for the Windows scheduler. Fix: bumped to +50ms (commit 0630fb5). Same class of bug has shipped at least three times in different files.
Two acceptable patterns. Anything else is a flake waiting to land on the worst possible PR.
Pattern A — fake timers (preferred when the test is about timing)
import { describe, it, expect, vi, beforeEach, afterEach } from 'vitest';
describe('cache TTL', () => {
beforeEach(() => vi.useFakeTimers());
afterEach(() => vi.useRealTimers());
it('expires after 5 minutes', () => {
const cache = makeCache({ ttlMs: 5 * 60 * 1000 });
cache.set('k', 'v');
vi.advanceTimersByTime(5 * 60 * 1000 + 1);
expect(cache.get('k')).toBeUndefined();
});
});
Pattern B — explicit tolerance (only when you must measure real wall-clock)
const start = Date.now();
await operationUnderTest();
const elapsed = Date.now() - start;
// Comment required — names the tolerance + the reason for it.
// +50ms accommodates the Windows scheduler; Phase 41 S5 used +5ms and flaked.
expect(elapsed).toBeGreaterThanOrEqual(target);
expect(elapsed).toBeLessThan(target + 50);
Banned patterns (caught by test-smells.mjs as TIME-FLAKE):
setTimeout(fn, N)in a test body withoutvi.useFakeTimers()Math.random()anywhere in a test — use a seeded RNG or fixtureDate.now()/new Date()withoutvi.setSystemTime()or an explicit tolerance commentperformance.now()without a tolerance comment
2. Never commit .only or focused tests
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- yesterday First seen · 230 lines · 3,216 tokens per session scan A d1ee182633e0
plan-forge testing.instructions.md is an instructions file published in the GitHub repository srnichols/plan-forge (5 stars, last pushed 21d ago), licensed MIT. It adds 3,216 tokens to every session, about $0.0161 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other instructions, from other repositories
dotnet-skills AGENTS.md
Instructions for managedcode/dotnet-skills, covering agents.md, purpose, solution topology, rule precedence and path and linking rules.
dotnet-skills copilot-instructions.md
Instructions for managedcode/dotnet-skills: Use AGENTS.md as the repository-wide source of truth for workflow, catalog structure, release policy, and skill maintenance rules.
Perigon.CLI copilot-instructions.md
Instructions for AterDev/Perigon.CLI, covering github copilot instructions, general guidelines, 技术栈, 项目结构与分层 and 代码风格约定.
copilot-instructions copilot-instructions.md
Instructions for SebastienDegodez/copilot-instructions, covering copilot instructions, language policy, development code generation and workflow implementation.
maf-doctor maf-deployment.instructions.md
Always-loaded production-deployment patterns for MAF 1.3.0. Auto-applies to Program.cs, DI registration files, and infra config. Covers ManagedIdentityCredential, MaxTokens caps, secret handling, OpenTelemetry wiring, and the analyzer rules that catch regressions at write time.
maf-doctor copilot-instructions.md
Instructions for joslat/maf-doctor, covering maf 1.3.0 migration — auto-loaded constraints, maf 1.3.0 — constraints & breaking changes reference, hard constraints (never violate), fan-out / fan-in rules (silent failure risk) and key breaking changes.