mcp-replay-dota2: Skill for Claude Code

.claude/skills/run-ci-and-test-replays/SKILL.md

run-ci-and-test-replays is a skill for Claude Code from DeepBlueCoding/mcp-replay-dota2. It costs 151 tokens per session (1,102 once invoked), scanned A, original, MIT.

A development guide for mcp-replay-dota2, an MCP server that works with Dota 2 match replay data. It defines the local checks and test fixtures used by the project.

In plain words
What is it for?
Run the project’s CI checks, write or repair replay-based tests, add fixtures, and test services using cached Dota 2 replay files.
Why use it?
It gives contributors a repeatable way to run formatting checks, type checks, and tests against consistent real-match data.

Skill for Claude Code

Written for Claude Code: installed under .claude/. Also seen: mentions CLAUDE.md.

This is DeepBlueCoding/mcp-replay-dota2's own configuration. It tells Claude Code how to work on mcp-replay-dota2 itself, so it is not a mod to install elsewhere. Copy it as a starting point and replace the rules that are about this project. Everything mcp-replay-dota2 configures →

Reuse

Borrowing it

Nothing to install: this file belongs to DeepBlueCoding/mcp-replay-dota2. Take a copy, put it at the same path in your own repository, and replace the rules that are about this project with yours.

Copy the file
curl -O https://raw.githubusercontent.com/DeepBlueCoding/mcp-replay-dota2/master/.claude/skills/run-ci-and-test-replays/SKILL.md
Clone the repo
git clone --depth 1 https://github.com/DeepBlueCoding/mcp-replay-dota2

Made for: Claude Code.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for run-ci-and-test-replays

README.md
[![agentmods](https://agentmods.dev/badge/skills/deepbluecoding/mcp-replay-dota2/run-ci-and-test-replays/github.svg)](https://agentmods.dev/skills/deepbluecoding/mcp-replay-dota2/run-ci-and-test-replays)
Your own site
<a href="https://agentmods.dev/skills/deepbluecoding/mcp-replay-dota2/run-ci-and-test-replays"><img src="https://agentmods.dev/badge/skills/deepbluecoding/mcp-replay-dota2/run-ci-and-test-replays/github.svg" alt="Measured on agentmods" height="20"></a>

Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.

agentmods 80×15 button for run-ci-and-test-replays

Your own site · 80×15
<a href="https://agentmods.dev/skills/deepbluecoding/mcp-replay-dota2/run-ci-and-test-replays"><img src="https://agentmods.dev/badge/skills/deepbluecoding/mcp-replay-dota2/run-ci-and-test-replays.svg" alt="Reviewed on agentmods" width="80" height="20"></a>
Per session 151 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 1,102 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00151 $0.01102
Opus 5 $0.00076 $0.00551
Sonnet 5 $0.00030 $0.00220
Haiku 4.5 $0.00015 $0.00110

Measured 10d ago against content hash 7bba691d2dee, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-11, from the pricing page.

Security

Grade A, and why

run-ci-and-test-replays scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 10d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

.claude/skills/run-ci-and-test-replays/SKILL.md · 85 lines

How it starts

The opening of the file, as written. The whole thing — 85 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Run CI and write replay tests for mcp-replay-dota2

CLAUDE.md (repo root) states the policy; this skill is the operational detail.

The CI gate (all three mandatory before any commit/push, in this order)

uv run ruff check src/ tests/ dota_match_mcp_server.py
uv run mypy src/ dota_match_mcp_server.py --ignore-missing-imports
uv run pytest

For one new test, run just it first:

uv run pytest tests/services/farming/test_farming_service.py::TestMultiCampDetection -v

Install deps with uv sync (or uv sync --group dev), NOT uv pip install. Always prefix with uv run. Never call python/pip directly.

Replay files & cache

Tests need two real replays in ~/dota2/replays/ (override the dir with DOTA_REPLAY_CACHE):

  • 8461956309.dem — primary, ~400 MB.
  • 8594217096.dem — secondary (OG match), smaller.

conftest.py parses each replay ONCE at session start (session-scoped, in-memory + diskcache) and injects ~65 pre-sliced fixtures (e.g. hero_deaths, combat_log_280_290, combat_log_280_290_earthshaker, objectives, rune_pickups, and the *_2 variants for match 2). asyncio_mode=auto, so async test functions need no decorator. Disk cache lives at ~/.cache/mcp_dota2/parsed_replays_v2 with a 7-day TTL.

NEVER call Parser or get_parsed_data directly in a test — request fixtures instead. If your test needs data no fixture provides, add a new session-scoped fixture in conftest.py that derives from the already-parsed data (don't parse again).

Golden-master verified constants (assert these EXACT values)

Match 8461956309: first blood = earthshaker killed by disruptor at 288.0s ("4:48").

Match 8594217096: first blood = batrider by pugna at 84.0s ("1:24"); total deaths after start = 53; roshan kills = 3; tower kills = 14; barracks kills = 6; rune pickups = 13; courier kills = 5.

Good test asserts a real victim/killer/time/count. Bad test asserts isinstance(...), len > 0, or key existence — those are banned (see CLAUDE.md).

Read the full file on GitHub · 85 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 10d ago First seen · 85 lines · 151 tokens per session scan A 7bba691d2dee

Subscribe to this mod's changes

run-ci-and-test-replays is a skill published in the GitHub repository DeepBlueCoding/mcp-replay-dota2 (2 stars, last pushed 3mo ago), licensed MIT. It adds 151 tokens to every session and 1,102 once invoked, about $0.0008 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other skills, from other repositories

Verification & Quality Assurance

Comprehensive truth scoring, code quality verification, and automatic rollback system with 0.95 accuracy threshold for ensuring high-quality agent outputs and codebase reliability.

ruvnet/ruflo · 36 tokens

ci

Configure Ginkgo for continuous integration — the recommended CLI flag set and the rationale for each flag (-r -p --randomize-all --randomize-suites --fail-on-pending --fail-on-empty --keep-going --cover --race --trace --json-report --timeout --poll-progress-after/-interval), invoking via go run to pin the CLI to…

onsi/ginkgo · 151 tokens

migrate-vstest-to-mtp

Use this skill before answering, planning, or editing whenever .NET tests or CI are switching from VSTest to Microsoft.Testing.Platform (MTP), or an MTP migration behaves differently. Triggers include "switch from VSTest"; MSTest/NUnit/xUnit MTP enablement; OutputType=Exe only for test projects in…

dotnet/skills · 191 tokens

playwright-ci

Production-ready CI/CD configurations for Playwright — GitHub Actions, GitLab CI, CircleCI, Azure DevOps, Jenkins, Docker, parallel sharding, reporting, code coverage, and global setup/teardown.

zebbern/claude-code-guide · 48 tokens

ci-maintenance-workflow

CI and GitHub Actions maintenance workflows — fix a failing test from a CI URL, fix a failing smoke test, add @pytest.mark.slow markers to slow tests, or review a PR against agent-checkable standards. Use when user asks to fix a failing test, fix a smoke test, mark slow tests, or review a PR. Trigger when the user…

UKGovernmentBEIS/inspect_evals · 120 tokens

playwright-testing

E2E testing with Playwright - Page Objects, cross-browser, CI/CD.

alinaqi/maggy · 20 tokens