devlab-test-expert

devlab-test-expert is a skill for Claude Code, Codex from seed-forge/harness-ai-kit. It costs 40 tokens per session (3,510 once invoked), scanned A, original, Apache-2.0.

A shared reference library for designing and troubleshooting software tests. It collects guidance on test strategy, test data, automation patterns, performance testing, flaky tests, accessibility, browser testing, and REST or GraphQL APIs.

In plain words
What is it for?
Use it when planning tests, writing acceptance criteria, managing test data, reviewing automation design, diagnosing flaky tests, checking accessibility, testing APIs, or evaluating AI services.
Why use it?
It gives other testing workflows a common source of practical rules and warnings about common mistakes. This helps teams choose suitable test levels and investigate failures consistently.

Skill for Claude CodeCodex

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/seed-forge/harness-ai-kit/devlab-test-expert
Any agent
npx skills add seed-forge/harness-ai-kit --skill devlab-test-expert
Clone the repo
git clone --depth 1 https://github.com/seed-forge/harness-ai-kit

Made for: Claude Code, Codex.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for devlab-test-expert

README.md
[![agentmods](https://agentmods.dev/badge/skills/seed-forge/harness-ai-kit/devlab-test-expert.svg)](https://agentmods.dev/skills/seed-forge/harness-ai-kit/devlab-test-expert)
Your own site
<a href="https://agentmods.dev/skills/seed-forge/harness-ai-kit/devlab-test-expert"><img src="https://agentmods.dev/badge/skills/seed-forge/harness-ai-kit/devlab-test-expert.svg" alt="Measured on agentmods" height="20"></a>
Per session 40 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 3,510 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00040 $0.03510
Opus 5 $0.00020 $0.01755
Sonnet 5 $0.00008 $0.00702
Haiku 4.5 $0.00004 $0.00351

Measured 4d ago against content hash ff70dfea4ec4, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

devlab-test-expert scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 4d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

skills/devlab-test-expert/SKILL.md · 387 lines

How it starts

The opening of the file, as written. The whole thing — 387 lines — stays where its author put it; the contents beside it link to each section on GitHub.

AI 驱动测试 - 最佳实践专家知识库

这是企业级测试团队的共享知识中枢。

聚合最佳实践、反模式预警、故障排除手册等专家知识,供所有测试子技能调用。


触发条件

当用户提到以下关键词时触发此 Skill:

  • "测试最佳实践" / "testing best practices"
  • "常见陷阱" / "antipatterns"
  • "验收标准怎么写" / "acceptance criteria examples"
  • "性能测试怎么做" / "performance testing guide"
  • "Flaky test 治理" / "flaky test tactics"

知识领域索引

📚 测试方法论

Reference 主题 适用场景
testing-maturity-model 测试能力成熟度模型 组织评估
shift-left-testing-strategies 左移测试策略 流程改进
test-data-management-best-practices 测试数据管理 数据准备
test-automation-pyramid 自动化测试金字塔 架构设计
property-based-testing 属性测试模式(fast-check / Hypothesis / jqwik) 随机数据、性质验证、最小反例
ai-service-test-tiering AI 服务测试分级隔离 marker 分级 + opt-in 开关 + 环境隔离

🎯 前端专项

Reference 主题 适用场景
ui-selection-stability 选择器稳定性指南 Web E2E
component-testing-patterns 组件测试模式 Vue/React
visual-regression-guide 视觉回归测试 UI 变更
a11y-testing-checklist 可访问性测试清单 WCAG 合规

🔧 后端/API 专项

Reference 主题 适用场景
api-testing-complete-guide API 测试完整指南 REST/GraphQL
authn-authz-testing 身份认证授权测试 安全测试
schema-validation-patterns Schema 验证模式 OpenAPI
rate-limiting-tests 限流降级测试 高可用

🛡️ 运维与质量

Reference 主题 适用场景
flaky-test-detection Flaky Test 检测与治理 CI/CD
test-performance-benchmarking 测试性能基准 性能工程
code-coverage-deep-dive 代码覆盖率深度分析 质量门禁
security-testing-integration 安全测试集成 AppSec

Read the full file on GitHub · 387 lines

Files

What ships with it

8 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 4d ago First seen · 387 lines · 40 tokens per session scan A ff70dfea4ec4

Subscribe to this mod's changes

devlab-test-expert is a skill published in the GitHub repository seed-forge/harness-ai-kit (21 stars, last pushed 3d ago), licensed Apache-2.0. It adds 40 tokens to every session and 3,510 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other skills, from other repositories

auth-web-cloudbase

CloudBase Web Authentication Quick Guide for frontend integration after auth-tool has already been checked. Provides concise and practical Web authentication solutions with multiple login methods and complete user management.

TencentCloudBase/CloudBase-AI-Toolkit · 38 tokens

browse-and-evaluate

Use when exploring the ai-agent-skills catalog to find, compare, and evaluate skills before installing. Always use --fields to limit output size and --dry-run before committing to an install.

MoizIbnYousaf/Ai-Agent-Skills · 43 tokens

loop-engineering

Shared loop-engineering reference for COG skills - the agent loop, deterministic verifiers, termination conditions, in-loop context management, and named patterns. Invoke when designing or debugging a skill that iterates (search-verify-retry, scan-until-dry, fetch-retry-gate).

huytieu/COG-second-brain · 63 tokens

telnyx-messaging-hosted-curl

Set up hosted SMS numbers, toll-free verification, and RCS messaging. Use when migrating numbers or enabling rich messaging features. This skill provides REST API (curl) examples.

team-telnyx/ai · 45 tokens

render-airdrop-carousel

Assemble a viral iOS "AirDrop" notification-carousel video ad (≈6–8s, 9:16) from a brand line plus 6–16 real product photos — a native AirDrop share-sheet card ("Brand would like to share a · Decline / Accept") springs up and its preview window CYCLES through the products, landing on a range/lineup payoff with an…

gooseworks-ai/goose-skills · 207 tokens

render-3d-product-showcase

Assemble a premium 3D product-showcase ad from a config — four beat clips (an orbiting hero rotation, a macro push-in, a physics reveal, a typographic close) normalized to the brand-color canvas, hard-concatenated in order, closed on a deterministic Playwright brand end card, and mixed under one instrumental bed at…

gooseworks-ai/goose-skills · 159 tokens