test-automation

test-automation is a skill for Claude Code, Codex from saolalab/clawforce. It costs 21 tokens per session (479 once invoked), scanned A, original, Apache-2.0.

A guide to automated software testing, covering unit, integration, end-to-end, performance, and API tests. It also explains test design practices and how to choose a testing framework.

In plain words
What is it for?
Use it when building, maintaining, or troubleshooting automated tests, choosing tools such as Jest, Pytest, Playwright, or Cypress, or planning a balanced test suite.
Why use it?
It helps teams select suitable tests and keep them fast, repeatable, and maintainable. It explains terms such as end-to-end testing, which checks a whole system rather than one small part.

Skill for Claude CodeCodex

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/saolalab/clawforce/test-automation
Any agent
npx skills add saolalab/clawforce --skill test-automation
Clone the repo
git clone --depth 1 https://github.com/saolalab/clawforce

Made for: Claude Code, Codex.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for test-automation

README.md
[![agentmods](https://agentmods.dev/badge/skills/saolalab/clawforce/test-automation.svg)](https://agentmods.dev/skills/saolalab/clawforce/test-automation)
Your own site
<a href="https://agentmods.dev/skills/saolalab/clawforce/test-automation"><img src="https://agentmods.dev/badge/skills/saolalab/clawforce/test-automation.svg" alt="Measured on agentmods" height="20"></a>
Per session 21 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 479 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00021 $0.00479
Opus 5 $0.00010 $0.00239
Sonnet 5 $0.00004 $0.00096
Haiku 4.5 $0.00002 $0.00048

Measured 5d ago against content hash 4f8ea9eb0550, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

test-automation scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 5d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

marketplace/roles/qa-engineer/workspace/skills/test-automation/SKILL.md · 87 lines

What it actually says

Test Automation

Test Pyramid

      /\
     /E2E\        <- Few, slow, expensive
    /------\
   /Integration\  <- Some, medium speed
  /--------------\
 /   Unit Tests   \ <- Many, fast, cheap
/------------------\

Test Types

Type Scope Speed Reliability
Unit Single function/class Fast High
Integration Multiple components Medium Medium
E2E Full system Slow Lower
Performance Load/stress Varies Medium

Automation Framework Selection

Considerations

  • Language compatibility
  • Community support
  • CI/CD integration
  • Reporting capabilities
  • Maintenance burden

Common Frameworks

  • Unit: Jest, Pytest, JUnit
  • Integration: Supertest, TestContainers
  • E2E: Playwright, Cypress, Selenium
  • API: Postman, REST Assured

Test Design Principles

FIRST

  • Fast — Tests run quickly
  • Independent — No dependencies between tests
  • Repeatable — Same result every time
  • Self-validating — Pass or fail, no interpretation
  • Timely — Written alongside code

AAA Pattern

def test_example():
    # Arrange - Set up test data
    user = create_test_user()
    
    # Act - Perform the action
    result = login(user.email, user.password)
    
    # Assert - Verify the outcome
    assert result.success == True

Flaky Test Management

Causes

  • Timing/race conditions
  • Shared state
  • External dependencies
  • Environment differences

Solutions

  • Add explicit waits
  • Isolate test data
  • Mock external services
  • Use test containers

Test Data Management

  • Factory patterns — Generate test data
  • Fixtures — Reusable setup
  • Seeding — Consistent database state
  • Cleanup — Reset after tests
Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 5d ago First seen · 87 lines · 21 tokens per session scan A 4f8ea9eb0550

Subscribe to this mod's changes

test-automation is a skill published in the GitHub repository saolalab/clawforce (38 stars, last pushed 4mo ago), licensed Apache-2.0. It adds 21 tokens to every session and 479 once invoked, about $0.0001 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other skills, from other repositories

agenticmail

🎀 AgenticMail — Full email, SMS, storage & multi-agent coordination for AI agents. 63 tools.

InternLM/WildClawBench · 28 tokens

agentic-paper-digest-skill

Fetches and summarizes recent arXiv and Hugging Face papers with Agentic Paper Digest. Use when the user wants a paper digest, a JSON feed of recent papers, or to run the arXiv/HF pipeline.

InternLM/WildClawBench · 54 tokens

ai-meeting-scheduling

Booking links fail for groups. SkipUp schedules meetings with 2-50 participants via email — one API call coordinates across timezones automatically. Also: check status, pause, resume, or cancel requests. Async only — does not instant-book, access calendars, or do free/busy lookups.

InternLM/WildClawBench · 66 tokens

arxiv-summarizer-orchestrator

End-to-end orchestration skill for periodic arXiv collection and reporting using three sub-skills: arxiv-search-collector, arxiv-paper-processor, and arxiv-batch-reporter. Supports manual language control across all markdown outputs and Stage-B processing strategy (subagentparallel default max 5, or serial).

InternLM/WildClawBench · 75 tokens

academic-literature-search

这是一个专注于学术文献检索的专业工具,集成了多个权威学术数据库,提供全面、快速、准确的文献检索服务。支持多数据库并发检索、高级过滤、智能排序和多种输出格式。.

InternLM/WildClawBench · 0 tokens

eachlabs-voice-audio

Text-to-speech, speech-to-text, voice conversion, and audio processing using EachLabs AI models. Supports ElevenLabs TTS, Whisper transcription with diarization, and RVC voice conversion. Use when the user needs TTS, transcription, or voice conversion.

InternLM/WildClawBench · 60 tokens