playwright-testing

A set of rules for using Playwright, a browser-automation tool, to test complete web user journeys. It supports tests across Chromium, Firefox, and WebKit browsers.

In plain words
What is it for?
Use it for end-to-end browser tests, API tests, screenshots, videos, parallel test runs, and HTML test reports.
Why use it?
It provides a governed way to run browser and API tests, capture evidence, and track reports while requiring human approval.

Skill for Claude CodeCodex

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/layer1labs/specsmith/playwright-testing
Any agent
npx skills add layer1labs/specsmith --skill playwright-testing
Clone the repo
git clone --depth 1 https://github.com/layer1labs/specsmith

Made for: Claude Code, Codex.

Per session 19 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 385 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00019 $0.00385
Opus 5 $0.00010 $0.00192
Sonnet 5 $0.00004 $0.00077
Haiku 4.5 $0.00002 $0.00038

Measured 2d ago against content hash 017db689ef4e, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

playwright-testing scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 2d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

.agents/skills/playwright-testing/SKILL.md · 73 lines

What it actually says

Playwright Testing Framework

This skill enables specsmith projects to use Playwright for end-to-end browser testing while maintaining governance compliance.

Requirements Covered

  • REQ-101: Playwright Testing Framework
  • Allows specsmith projects to use Playwright for end-to-end browser testing with human approval

Integration Capabilities

Framework Features

  • End-to-end browser testing with Playwright
  • Cross-browser testing (Chromium, Firefox, WebKit)
  • API testing capabilities
  • Screenshot and video capture
  • Parallel test execution
  • Test reporting and analytics

Governance Requirements

  • Human approval required for all Playwright testing
  • Configuration management with governance controls
  • Test artifact management with approval
  • Reporting and coverage tracking with approval

Usage Examples

# Initialize Playwright testing
specsmith skill install playwright-testing

# Configure Playwright settings
specsmith playwright configure --browser chromium --timeout 30000

# Run end-to-end tests
specsmith playwright run --test-suite e2e-tests

# Generate test reports
specsmith playwright report --format html

Configuration

The skill requires the following configuration in scaffold.yml:

playwright:
  browser: chromium
  timeout: 30000
  headless: true
  parallel: true
  screenshot: "on_failure"

Security Considerations

  • All Playwright testing requires human approval
  • Test environment isolation must be maintained
  • Sensitive data handling in tests must be secured
  • Test artifact access controls must be enforced

Compliance Requirements

  • Follow security best practices for testing environments
  • Maintain audit logs of all Playwright test executions
  • Ensure proper test artifact management and retention
  • Generate and track test coverage reports
Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 2d ago First seen · 73 lines · 19 tokens per session scan A 017db689ef4e

Subscribe to this mod's changes

playwright-testing is a skill published in the GitHub repository layer1labs/specsmith (7 stars, last pushed 29d ago), licensed MIT. It adds 19 tokens to every session and 385 once invoked, about $0.0001 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other skills, from other repositories

accessibility-audit

Unified accessibility audit - axe-core WCAG 2.2 scan, APCA contrast measurement, static a11y checks (aria-label, tabindex, form labels, focus visibility, onClick on non-interactive). Static and live modes.

marcoguillermaz/Tierward · 53 tokens

arch-audit

Audit Claude Code architecture files against Anthropic docs and release notes, and verify internal ecosystem consistency. Run on demand to maintain compliance, catch new features, and keep the context system clean.

marcoguillermaz/Tierward · 41 tokens

pr-review

Autonomous local code review. Reviews either an open PR's diff (via gh) or the local working diff against a base branch (--local, auto-selected when no PR exists yet), spawns a dedicated review subagent with universal + stack-specific severity criteria, posts the review as a comment on the PR (audit trail; skipped in…

marcoguillermaz/Tierward · 167 tokens

security-audit

Security audit: auth/authz on API routes, input validation, row-level access control (RLS in Postgres, ORM scopes or app-level guards elsewhere), response shape review, secret exposure, HTTP headers. Native mode checks entitlements and Keychain. MCP-aware (v1.20+): when mcp-nvd server is wired, Step 3c queries live…

marcoguillermaz/Tierward · 109 tokens

skill-db

Database audit: schema quality, index coverage, row-level access-control completeness, FK cascades, query patterns. Runs live SQL verification (PostgreSQL instance in PATTERNS.md; other engines verify the equivalent guard). Migration file safety → /migration-audit.

marcoguillermaz/Tierward · 56 tokens

skill-dev

Code quality audit: detect cross-module coupling, N+1 queries, dead exports, antipatterns, over-large components. Cross-checks against refactoring backlog.

marcoguillermaz/Tierward · 36 tokens