Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx skills add mattmre/EVOKORE-MCP-PUBLIC --skill orch-verifygit clone --depth 1 https://github.com/mattmre/EVOKORE-MCP-PUBLICWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/mattmre/evokore-mcp-public/orch-verify)<a href="https://agentmods.dev/skills/mattmre/evokore-mcp-public/orch-verify"><img src="https://agentmods.dev/badge/skills/mattmre/evokore-mcp-public/orch-verify/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/skills/mattmre/evokore-mcp-public/orch-verify"><img src="https://agentmods.dev/badge/skills/mattmre/evokore-mcp-public/orch-verify.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00017 | $0.00575 |
| Opus 5 | $0.00009 | $0.00287 |
| Sonnet 5 | $0.00003 | $0.00115 |
| Haiku 4.5 | $0.00002 | $0.00057 |
Grade A, and why
orch-verify scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 8d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 110 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Orchestration Verify
Purpose
Capture verification evidence for the current task by running tests, capturing output, and updating the verification log.
Invocation
orch-verify [command]
Arguments
| Argument | Required | Description |
|---|---|---|
| command | No | Specific test command to run |
If no command is provided, uses the default test command from task context or project configuration.
Workflow
1. Context Load
Read the following:
- Current task from task queue (status:
in_progress) - Test commands from project configuration or status tracking
- Evidence capture protocol (if defined)
2. Test Execution
Execute verification command(s):
- Run specified command or default test suite
- Capture stdout, stderr, and exit code
- Record execution timestamp
3. Evidence Capture
Collect verification artifacts:
- Command: Exact command executed
- Exit Code: Process return code
- Output: Relevant output (truncated if excessive)
- Timestamp: ISO 8601 execution time
- Task Reference: Associated task ID
4. Log Update
Append entry to verification-log.md:
## [YYYY-MM-DD HH:MM] Task: [task-id]
**Command**:
[executed command]
**Result**: [PASS | FAIL]
**Exit Code**: [code]
**Output**:
[captured output, truncated to 50 lines]
**Evidence Hash**: [sha256 of output, optional]
---
5. Status Update
If verification passes:
- Update task status per workflow (may advance to
doneorreview) - Report success summary
If verification fails:
- Keep task status unchanged
- Report failure with relevant output excerpt
Outputs
| Output | Destination | Action |
|---|---|---|
| Verification result | stdout | display |
| Evidence entry | verification-log.md |
append |
Evidence Standards
- All verification must be captured with reproducible commands
- Output should be sufficient to demonstrate pass/fail
- Sensitive data must be redacted before logging
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 8d ago First seen · 110 lines · 17 tokens per session scan A 78f9a74ab21e
orch-verify is a skill published in the GitHub repository mattmre/EVOKORE-MCP-PUBLIC (3 stars, last pushed 3mo ago), licensed MIT. It adds 17 tokens to every session and 575 once invoked, about $0.0001 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-09-03.
Other skills, from other repositories
design-qa-checklist
Build a QA checklist for verifying that a build matches the design. Use at implementation review. For the spec engineers build from, use handoff-spec.
testing
Use this skill when writing, reviewing, or improving tests in WrongStack. Triggers: user says "test", "unit test", "integration test", "e2e", "mock", "vitest", "coverage", "assert", "expect", "test strategy", "write tests".
autoresearch
Autonomously optimize any Claude Code skill by running it repeatedly, scoring outputs against binary evals, mutating the prompt, and keeping improvements. Based on Karpathy's autoresearch methodology. Use when: optimize this skill, improve this skill, run autoresearch on, make this skill better, self-improve skill…
hunt-race-condition
Hunting skill for race condition vulnerabilities. Built from 3 public bug bounty reports. Use when hunting race condition on any target.
browserstack
Run tests on BrowserStack. Use when user mentions "browserstack", "cross-browser", "cloud testing", "browser matrix", "test on safari", "test on firefox", or "browser compatibility".
generate
Generate Playwright tests. Use when user says "write tests", "generate tests", "add tests for", "test this component", "e2e test", "create test for", "test this page", or "test this feature".