plan-forge testing.instructions.md

Testing instructions for Plan Forge projects that use Vitest, a JavaScript and TypeScript testing tool. They cover time-dependent tests, mocking, fake timers, and how to interpret test output.

In plain words
What is it for?
Use them when editing, debugging, or reporting on tests, including choosing fake timers or a suitable timing tolerance and checking for common test problems.
Why use it?
They reduce flaky tests, especially tests that depend on real clock timing or vary across operating systems.

Instructions file for GitHub Copilot

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add instructions/srnichols/plan-forge/testing
Clone the repo
git clone --depth 1 https://github.com/srnichols/plan-forge

Made for: GitHub Copilot.

Per session 3,216 This file is loaded in full into every session.
When invoked 3,216 The same file — it is already loaded in full.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.03216 $0.03216
Opus 5 $0.01608 $0.01608
Sonnet 5 $0.00643 $0.00643
Haiku 4.5 $0.00322 $0.00322

Measured yesterday against content hash d1ee182633e0, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

plan-forge testing.instructions.md scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured yesterday.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

.github/instructions/testing.instructions.md · 230 lines

How it starts

The opening of the file, as written. The whole thing — 230 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Testing Instructions

When this loads: every time you edit, write, debug, or report on a test file. Sister script: scripts/audit/test-smells.mjs — mechanical scan for the patterns this file rules against.


The 6 Rules

1. Time-sensitive tests must declare tolerance OR use fake timers

The single biggest source of CI flake. Phase 41 Slice 5 reference incident: timeline-core cache-invalidation test used a +5ms tolerance — too tight for the Windows scheduler. Fix: bumped to +50ms (commit 0630fb5). Same class of bug has shipped at least three times in different files.

Two acceptable patterns. Anything else is a flake waiting to land on the worst possible PR.

Pattern A — fake timers (preferred when the test is about timing)

import { describe, it, expect, vi, beforeEach, afterEach } from 'vitest';

describe('cache TTL', () => {
  beforeEach(() => vi.useFakeTimers());
  afterEach(() => vi.useRealTimers());

  it('expires after 5 minutes', () => {
    const cache = makeCache({ ttlMs: 5 * 60 * 1000 });
    cache.set('k', 'v');
    vi.advanceTimersByTime(5 * 60 * 1000 + 1);
    expect(cache.get('k')).toBeUndefined();
  });
});

Pattern B — explicit tolerance (only when you must measure real wall-clock)

const start = Date.now();
await operationUnderTest();
const elapsed = Date.now() - start;
// Comment required — names the tolerance + the reason for it.
// +50ms accommodates the Windows scheduler; Phase 41 S5 used +5ms and flaked.
expect(elapsed).toBeGreaterThanOrEqual(target);
expect(elapsed).toBeLessThan(target + 50);

Banned patterns (caught by test-smells.mjs as TIME-FLAKE):

  • setTimeout(fn, N) in a test body without vi.useFakeTimers()
  • Math.random() anywhere in a test — use a seeded RNG or fixture
  • Date.now() / new Date() without vi.setSystemTime() or an explicit tolerance comment
  • performance.now() without a tolerance comment

2. Never commit .only or focused tests

Read the full file on GitHub · 230 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. yesterday First seen · 230 lines · 3,216 tokens per session scan A d1ee182633e0

Subscribe to this mod's changes

plan-forge testing.instructions.md is an instructions file published in the GitHub repository srnichols/plan-forge (5 stars, last pushed 21d ago), licensed MIT. It adds 3,216 tokens to every session, about $0.0161 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other instructions, from other repositories

dotnet-skills AGENTS.md

Instructions for managedcode/dotnet-skills, covering agents.md, purpose, solution topology, rule precedence and path and linking rules.

managedcode/dotnet-skills · 13,592 tokens

dotnet-skills copilot-instructions.md

Instructions for managedcode/dotnet-skills: Use AGENTS.md as the repository-wide source of truth for workflow, catalog structure, release policy, and skill maintenance rules.

managedcode/dotnet-skills · 97 tokens

Perigon.CLI copilot-instructions.md

Instructions for AterDev/Perigon.CLI, covering github copilot instructions, general guidelines, 技术栈, 项目结构与分层 and 代码风格约定.

AterDev/Perigon.CLI · 1,254 tokens

copilot-instructions copilot-instructions.md

Instructions for SebastienDegodez/copilot-instructions, covering copilot instructions, language policy, development code generation and workflow implementation.

SebastienDegodez/copilot-instructions · 364 tokens

maf-doctor maf-deployment.instructions.md

Always-loaded production-deployment patterns for MAF 1.3.0. Auto-applies to Program.cs, DI registration files, and infra config. Covers ManagedIdentityCredential, MaxTokens caps, secret handling, OpenTelemetry wiring, and the analyzer rules that catch regressions at write time.

joslat/maf-doctor · 1,884 tokens

maf-doctor copilot-instructions.md

Instructions for joslat/maf-doctor, covering maf 1.3.0 migration — auto-loaded constraints, maf 1.3.0 — constraints & breaking changes reference, hard constraints (never violate), fan-out / fan-in rules (silent failure risk) and key breaking changes.

joslat/maf-doctor · 1,604 tokens