copilotd-e2e-verification

copilotd-e2e-verification is a skill for Claude Code, Codex from DamianEdwards/copilotd. It costs 31 tokens per session (1,756 once invoked), scanned A, original, MIT.

A testing guide for checking copilotd, a service that coordinates coding-agent work through GitHub issues and pull requests. It tests the complete process with real GitHub records, sessions, worktrees, and saved state.

In plain words
What is it for?
Use it to validate major changes to copilotd’s rules, sessions, worktrees, prompts, or GitHub issue and pull-request feedback loops.
Why use it?
Build checks and small command tests may miss problems in how the service coordinates work over time. This helps reveal broken dispatch, callbacks, cleanup, or state updates.

Skill for Claude CodeCodex

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/damianedwards/copilotd/copilotd-e2e-verification
Any agent
npx skills add DamianEdwards/copilotd --skill copilotd-e2e-verification
Clone the repo
git clone --depth 1 https://github.com/DamianEdwards/copilotd

Made for: Claude Code, Codex.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for copilotd-e2e-verification

README.md
[![agentmods](https://agentmods.dev/badge/skills/damianedwards/copilotd/copilotd-e2e-verification.svg)](https://agentmods.dev/skills/damianedwards/copilotd/copilotd-e2e-verification)
Your own site
<a href="https://agentmods.dev/skills/damianedwards/copilotd/copilotd-e2e-verification"><img src="https://agentmods.dev/badge/skills/damianedwards/copilotd/copilotd-e2e-verification.svg" alt="Measured on agentmods" height="20"></a>
Per session 31 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 1,756 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00031 $0.01756
Opus 5 $0.00015 $0.00878
Sonnet 5 $0.00006 $0.00351
Haiku 4.5 $0.00003 $0.00176

Measured 4d ago against content hash 83cea94d48cc, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

copilotd-e2e-verification scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 4d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

.github/skills/copilotd-e2e-verification/SKILL.md · 103 lines

How it starts

The opening of the file, as written. The whole thing — 103 lines — stays where its author put it; the contents beside it link to each section on GitHub.

copilotd end-to-end verification

Use this skill when validating substantial copilotd orchestration changes, especially changes to dispatch rules, reconciliation, session lifecycle, worktree handling, prompt callbacks, or GitHub issue/PR feedback loops.

Goal

Verify copilotd as a live reconciliation daemon, not just by build or command smoke tests. A good verification proves that real GitHub issues and pull requests move through the expected lifecycle, Copilot sessions can call back into copilotd, worktrees and state converge correctly, and temporary artifacts are cleaned up.

Optimum approach

  1. Use an isolated copilotd home. Set COPILOTD_HOME to a disposable directory such as .\copilotd-home-e2e or another task-specific path. Do not use the user's normal ~\.copilotd state.
  2. Start with local CLI validation. Build the branch, then verify config, rules add/list/update/delete, invalid rule arguments, status, session list --all, and JSON persistence before creating live artifacts.
  3. Use a real but safe target repository. Prefer a repo already cloned locally and writable by the user. Configure repo_home so copilotd resolves the existing clone and creates sibling <repo>_sessions worktrees.
  4. Create explicit temporary labels. Use unique labels such as copilotd-e2e-ready, copilotd-e2e-clarify, and copilotd-e2e-pr so rules match only the verification artifacts.
  5. Create at least two issue scenarios.
    • A ready-to-implement issue with small, deterministic acceptance criteria that should lead to a branch, commit, PR, session pr, and WaitingForReview.
    • An intentionally ambiguous issue that should lead to session comment, WaitingForFeedback, a real clarification reply, and re-dispatch.
  6. Create a PR rule that observes the issue-created PRs. Add a PR dispatch rule using --kind pr, a temporary PR label, --base, and a safe branch strategy such as read-only for validation-only sessions.
  7. Drive the full feedback loop. After issue sessions create PRs, let the PR rule launch PR-root validation sessions. Ensure those sessions comment on the PR. Then add a manual PR comment or review from a trusted user without a copilotd marker so the original issue-owned WaitingForReview session re-dispatches, pushes a follow-up commit, and returns to WaitingForReview.
  8. Observe state and GitHub together. Cross-check state.json, copilotd session list --all, issue comments, PR comments, PR labels, PR commits, and local git worktrees. Do not trust a single surface.
  9. Keep a journal. Record every issue, workaround, and stop-worthy finding as it happens, including exact issue/PR numbers and whether the behavior was expected or a product problem.
  10. Clean up aggressively. Close temporary PRs without merging, close temporary issues, delete temporary branches/labels, stop daemon processes, remove isolated homes and temporary publish folders, prune worktrees, and verify the target repo returns to a clean default branch.

Read the full file on GitHub · 103 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 4d ago First seen · 103 lines · 31 tokens per session scan A 83cea94d48cc

Subscribe to this mod's changes

copilotd-e2e-verification is a skill published in the GitHub repository DamianEdwards/copilotd (59 stars, last pushed 7d ago), licensed MIT. It adds 31 tokens to every session and 1,756 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other skills, from other repositories

winui-ui-testing

Automated UI testing for Windows desktop apps — generate a batch test script with the winapp ui UI Automation harness, run all tests in one pass, read results. Covers element assertions, interactions, value checking (TextBox, ComboBox, ToggleSwitch), keyboard shortcuts and typing (send-keys), hover, drag-and-drop…

microsoft/win-dev-skills · 122 tokens

pr-review

Multi-dimensional review of a PR or feature branch in microsoft/win-dev-skills. Activate on "review my PR / changes / branch", "vet before pushing", "PR review", "is this ready to merge". Fans out parallel sub-agents over skill content, the skill-vs-tool boundary (solution hierarchy), tool correctness…

microsoft/win-dev-skills · 98 tokens

winui-design

Use when designing, reviewing, or fixing WinUI 3: sample and control discovery with winapp find-ui, layout planning, control choice, Fluent Design alignment, Light/Dark/High Contrast theming, typography, spacing, brushes, accessibility, and XAML data-binding design. Load before authoring new XAML, reviewing UI PRs…

microsoft/win-dev-skills · 119 tokens

winui-dev-workflow

Build and run workflow for WinUI 3 apps with WinApp CLI 0.6+ — project creation with winapp new, project-mode winapp run, BuildAndRun.ps1 analyzer integration, crash diagnosis, and prerequisites. Use when creating, building, running, or fixing build errors in a WinUI 3 project.

microsoft/win-dev-skills · 73 tokens

winui-packaging

MSIX packaging, code signing, and distribution for WinUI 3 apps — build for release, certificate generation (winapp cert generate), certificate trust, code signing (winapp sign), self-contained deployment, CI/CD with GitHub Actions, and Microsoft Store submission. Use when preparing for release, creating MSIX…

microsoft/win-dev-skills · 86 tokens

winui-session-report

Analyze the current or a recent agent session (GitHub Copilot CLI or Claude Code) and generate a diagnostic report. Use only when the user explicitly asks for session feedback, agent debugging, or a review of what happened during a build session. Do not inspect session data automatically.

microsoft/win-dev-skills · 61 tokens