e2e-tests-reviewer

e2e-tests-reviewer is an agent for coding agents from vladolaru/claude-code-plugins. It costs 39 tokens per session (1,014 once invoked), scanned B, original, MIT.

A code review agent for browser end-to-end tests written with Playwright, a tool that drives a real browser to test complete user flows. It focuses on locators, page objects, waiting, network handling, and WordPress or WooCommerce test helpers.

In plain words
What is it for?
Use it to review browser-test selectors, Page Object Model code, automatic waiting, network interception, test setup, and WordPress or WooCommerce testing patterns.
Why use it?
It helps expose tests that pass for the wrong reasons, break easily when the interface changes, or fail to isolate each test.

Agent

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add agents/vladolaru/claude-code-plugins/e2e-tests-reviewer
Clone the repo
git clone --depth 1 https://github.com/vladolaru/claude-code-plugins

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for e2e-tests-reviewer

README.md
[![agentmods](https://agentmods.dev/badge/agents/vladolaru/claude-code-plugins/e2e-tests-reviewer.svg)](https://agentmods.dev/agents/vladolaru/claude-code-plugins/e2e-tests-reviewer)
Your own site
<a href="https://agentmods.dev/agents/vladolaru/claude-code-plugins/e2e-tests-reviewer"><img src="https://agentmods.dev/badge/agents/vladolaru/claude-code-plugins/e2e-tests-reviewer.svg" alt="Measured on agentmods" height="20"></a>
Per session 39 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 1,014 The whole file, excluding the scripts and references it only reads on demand.
Security scan B 1 finding. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00039 $0.01014
Opus 5 $0.00019 $0.00507
Sonnet 5 $0.00008 $0.00203
Haiku 4.5 $0.00004 $0.00101

Measured 4d ago against content hash 63a07deefa3e, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade B, and why

e2e-tests-reviewer scanned grade B with 1 finding against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 4d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Asks the agent to reveal its instructionsmediumSystem prompt leakage

Directions to print, repeat or translate the system prompt extract configuration the operator did not intend to expose.

Read the output carefully. It contains your review rules (including the shared tests protocol), review scope, and output instructions. If STATUS is NO_DOMAIN_FILES, report "No E2E test files to review" → APPROVE → exit.
plugins/pirategoat-tools/agents/e2e-tests-reviewer.md · 83 lines

How it starts

The opening of the file, as written. The whole thing — 83 lines — stays where its author put it; the contents beside it link to each section on GitHub.

MANDATORY SETUP — Run Bootstrap Before Reviewing

Do NOT start reviewing code until this step is done:

Run the bootstrap script:

PLUGIN_ROOT=$(cat /tmp/.pirategoat-tools-root 2>/dev/null)
[ -z "$PLUGIN_ROOT" ] || [ ! -d "$PLUGIN_ROOT/scripts" ] && PLUGIN_ROOT=$(find ~/.claude -path "*/pirategoat-tools/*/scripts/review/agent/bootstrap.py" -type f 2>/dev/null | sort | tail -1 | xargs dirname | xargs dirname | xargs dirname | xargs dirname)
python3 $PLUGIN_ROOT/scripts/review/agent/bootstrap.py --agent e2e-tests-reviewer

Read the output carefully. It contains your review rules (including the shared tests protocol), review scope, and output instructions. If STATUS is NO_DOMAIN_FILES, report "No E2E test files to review" → APPROVE → exit. If ERROR, follow the instructions and exit.


You are an expert E2E Test Quality Reviewer specializing in Playwright, Page Object Model, and WordPress/WooCommerce end-to-end testing.

Your expertise: Playwright locator strategies, auto-waiting patterns, Page Object Model architecture, network interception, test isolation, and E2E-specific anti-patterns.

This review matters. False confidence from bad tests causes production bugs that proper review would have caught.

Core Mission

Verify test resilience -> Detect flaky patterns -> Ensure isolation

Do NOT review implementation code. Do NOT review PHP unit tests or Jest/Vitest unit tests.

Deep Knowledge References

All reference files are at $PLUGIN_ROOT/skills/testing-patterns/references/.

Test Issue Reference File Sections to Read
Behavior vs implementation test-philosophy.md ## The Fundamental Shift + ## Four Core Principles
Flaky/brittle/slow tests test-smells.md ## The Six Major Test Smells (relevant subsection)
Mock usage decisions mocking-strategies.md ## The Mocking Decision Framework + ## Types of Test Doubles
E2E/Playwright patterns playwright-patterns.md Full file (~461L, manageable)

Read the full file on GitHub · 83 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 4d ago First seen · 83 lines · 39 tokens per session scan B 63a07deefa3e

Subscribe to this mod's changes

e2e-tests-reviewer is an agent published in the GitHub repository vladolaru/claude-code-plugins (8 stars, last pushed today), licensed MIT. It adds 39 tokens to every session and 1,014 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it B with 1 finding (asks the agent to reveal its instructions). No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other agents, from other repositories

trellis-check

Code quality check expert. Reviews code changes against specs and self-fixes issues.

yuqie6/ProductFlow · 20 tokens

strategy-consultant

You are a management and startup consultant for Korean founders, small-business owners, and startup operators. You turn a goal (validate business idea X, size market Y, win grant program Z, assess this storefront location) into concrete, evidence-based deliverables: business plans, business model canvases, market…

modu-ai/moai-cowork · 106 tokens

hemmabo-federation

Use when working with the HemmaBo federation MCP server: understanding pricing logic, checking availability rules, reviewing booking flows, explaining the ACP/Stripe integration, or making code changes to the server. Trigger phrases: hemmabo, federation, MCP server, pricing, booking, availability, ACP, villaakerlyckan.

HemmaBo-se/hemmabo-mcp-server · 64 tokens

shopify-app-architect

Use when starting a new Shopify app or designing a major feature. Specializes in creating complete architecture plans including data models, API routes, webhooks, scopes, billing strategy, and deployment targets. Route here for architecture approval workflows.

khadinakbarlabs/shopify-app-builder · 52 tokens

business-architect

Senior business-domain architect for SaaS, ERP, e-commerce, and full-stack applications. Delegates here for designing business logic that survives real-world edge cases — billing, multi-tenancy, inventory, GL postings, refunds, RBAC, audit, idempotency. Knows how Stripe, Linear, NetSuite, Shopify, and similar solve…

viknesh20-20/claude-code-tool-kit · 76 tokens

comps-researcher

Least-privilege pricing-research subagent. Pulls live eBay sold comps (and optional FB Marketplace context) for one secondhand item and returns exactly two labeled Markdown blocks. Use when the listing orchestrator needs pricing without the browser-heavy page text polluting the main context. Scraper only — never…

hansohn/marketplace-skills · 74 tokens