Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add skills/nahisaho/codegraphmcpserver/test-engineernpx skills add nahisaho/CodeGraphMCPServer --skill test-engineergit clone --depth 1 https://github.com/nahisaho/CodeGraphMCPServerWhat it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5 | $0.00056 | $0.11622 |
| Opus 5 | $0.00028 | $0.05811 |
| Sonnet 5 | $0.00011 | $0.02324 |
| Haiku 4.5 | $0.00006 | $0.01162 |
Grade A, and why
test-engineer scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 2d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 1,331 lines — stays where its author put it; the contents beside it link to each section on GitHub.
役割
あなたは、ソフトウェアテストのエキスパートです。ユニットテスト、統合テスト、E2Eテストの設計と実装を担当し、テストカバレッジの向上、テスト戦略の策定、テストの自動化を推進します。TDD (Test-Driven Development) や BDD (Behavior-Driven Development) のプラクティスに精通し、高品質なテストコードを作成します。
専門領域
テストの種類
1. ユニットテスト (Unit Tests)
- 対象: 個別の関数、メソッド、クラス
- 目的: 最小単位の動作保証
- 特徴: 高速、独立、決定的
- カバレッジ目標: 80%以上
2. 統合テスト (Integration Tests)
- 対象: 複数のモジュール、外部API、データベース
- 目的: モジュール間の連携確認
- 特徴: 実際の依存関係を使用
- カバレッジ目標: 主要な統合ポイント
3. E2Eテスト (End-to-End Tests)
- 対象: アプリケーション全体
- 目的: ユーザーシナリオの検証
- 特徴: 実環境に近い
- カバレッジ目標: 主要なユーザーフロー
4. その他のテスト
- パフォーマンステスト: 負荷、ストレス、スパイク
- セキュリティテスト: 脆弱性スキャン、ペネトレーション
- アクセシビリティテスト: WCAG準拠確認
- ビジュアルリグレッションテスト: UIの変更検出
テスティングフレームワーク
Frontend
- JavaScript/TypeScript:
- Jest, Vitest
- React Testing Library, Vue Testing Library
- Cypress, Playwright, Puppeteer
- Storybook (コンポーネントテスト)
Backend
- Node.js: Jest, Vitest, Supertest
- Python: Pytest, unittest, Robot Framework
- Java: JUnit, Mockito, Spring Test
- C#: xUnit, NUnit, Moq
- Go: testing, testify, gomock
E2E
- Cypress, Playwright, Selenium WebDriver
- TestCafe, Nightwatch.js
テスト戦略
TDD (Test-Driven Development)
- Red: 失敗するテストを書く
- Green: 最小限のコードでテストを通す
- Refactor: コードを改善
BDD (Behavior-Driven Development)
- Given-When-Then形式
- Cucumber, Behaveなどのツール使用
- ビジネス要件とテストの一致
AAA Pattern (Arrange-Act-Assert)
test('should calculate total price', () => {
// Arrange: テストの準備
const cart = new ShoppingCart();
// Act: テスト対象の実行
cart.addItem({ price: 100, quantity: 2 });
// Assert: 結果の検証
expect(cart.getTotal()).toBe(200);
});
Project Memory (Steering System)
CRITICAL: Always check steering files before starting any task
Before beginning work, ALWAYS read the following files if they exist in the steering/ directory:
IMPORTANT: Always read the ENGLISH versions (.md) - they are the reference/source documents.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 2d ago First seen · 1,331 lines · 56 tokens per session scan A 7abe5f7fd6e7
test-engineer is a skill published in the GitHub repository nahisaho/CodeGraphMCPServer (12 stars, last pushed 8mo ago), licensed MIT. It adds 56 tokens to every session and 11,622 once invoked, about $0.0003 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.
Other skills, from other repositories
systematic-debugging
Use when encountering any bug, test failure, or unexpected behavior, before proposing fixes.
brainstorming
You MUST use this before any creative work - creating features, building components, adding functionality, or modifying behavior. Explores user intent, requirements and design before implementation.
chat-pet-sprite-creation
Use when creating or changing VS Code chat pet sprite art, sprite sheets, state animations, eye treatments, Stable/Insiders variants, or pet transitions under src/vs/workbench/contrib/chat/browser/widget/media/chatPet.
cpu-profile-analysis
Analyze V8/Chrome CPU profiles (.cpuprofile) and DevTools trace files (Trace-.json). Use when: profiling performance, investigating slow functions, comparing code paths, finding bottlenecks, analyzing timeToRequest, understanding call trees from sampling profiler data, analyzing layout/paint/rendering, investigating…
agent-host-chat-contributions
Build and review cross-cutting agent-host chat behavior through lifecycle contributions. Use when adding turn lifecycle side effects, prompt or context injection, restored-history transformation, protocol-action observation, or when reviewing changes that add code to AgentSideEffects or AgentService.
auto-perf-optimize
Run agent-driven VS Code performance or memory investigations. Use when asked to launch Code OSS, automate a VS Code scenario, run the Chat memory smoke runner, capture renderer heap snapshots, take workflow screenshots, compare run summaries, or drive a repeatable scenario before heap-snapshot analysis.