Testet alle Kanban-Tasks im Status "Review" automatisch mit Chrome/Playwright und VHS (Terminal). Erkennt automatisch ob ein Task Browser- oder Terminal-Tests braucht. Liest Task-Notes, startet Server wenn nötig, testet visuell und funktional, und dokumentiert Ergebnisse in den Task-Notes. Nutze diesen Skill wenn JoPa…
Esta skill transforma al agente en un Ingeniero de QA especializado en Yari Medic. Su misión es auditar código contra los documentos de grounding en data/projectcontext/.
Parent Agent uses playwright-tabbed to allocate browser tabs and partition tasks, then spawns child agents in parallel — each operating on its assigned tabid — for browser acceptance testing or automation, with a final aggregated summary. Suitable for multi-page validation, parallel module checks, and…
A browser automation testing skill combining Playwright's precise controls with browser-use's AI-guided exploration. It operates a shared Chromium browser session.
Verify and debug your own web UI on localhost. Use when you changed frontend code and need to check it actually works — console/network errors, layout/contrast/a11y problems, dead buttons, wrong colors, broken handlers — or when testing a dev server, reproducing a UI bug, or debugging why an element behaves wrong.…
Write the behaviour specification for a feature BEFORE implementing it, as test cases. Use whenever you are about to build a new endpoint, screen, flow, or behaviour — especially anything touching auth, sessions, tokens, lists, or pagination. Also use when asked "how should I test this", "what should this actually…
Compares JSON actual vs expected template via the validate-json MCP tool. Use when validating API responses, JSON configs, nested objects, JSON Schema contracts, or when the user asks to compare, diff, or validate JSON data.
Security QA tester workflow for the owasp-wstg MCP server — scope, search, neighbor-walk, and citation over the WSTG corpus using five MCP tools. Load this file first; then load the model-variant file that matches your model family.
Generic, service-agnostic API test scenario design skill. Consumes an API contract from a Documentation MCP (or any authoritative contract source) and produces traceable, prioritized, coverage-reported test scenarios. Works for any REST/HTTP endpoint of any service/version — no hardcoded API knowledge.
A skill for running step-by-step tests on a real Android phone, using screen-reading and device-control tools while checking the result after each action.
Automated desktop UI testing through the Hyprland MCP server. Captures screenshots, analyzes them via the vision skill, clicks and types, and re-verifies. Use when asked to test a desktop app, verify UI behavior, run a visual regression check, or drive a background app without interrupting the user. Triggers: "test…
Use when running, extending, or diagnosing batch tests against the Telegraph miner or chatbot API, including JSON query validation, resumable JSONL collection, request failures, and coverage summaries.
Generate local test fixtures via Test Files MCP. Use when the user needs dummy files, images, PDFs, CSVs, JSON, ZIPs, WAV audio, corrupted files, or edge-case filenames for upload testing, QA, parser testing, or validation.
Guide for building a Phoenix-based evaluation harness for an MCP server. Use when you need to create or extend LLM-as-judge evaluations that measure how well an agent can accomplish realistic tasks using only the MCP server's tools. Covers the eval/phoenix module layout, dataset creation, the LangChain agent, judges…
Run the eval harness over 50+ synthetic COD orders and produce batch metrics with honest assumptions. Use when validating gate decisions or preparing evidence for the buildathon submission.
Drive a VUnit (HDL unit-testing) project through the vunit-mcp MCP server (vunitstatus, vunitlisttests, vunitlistfiles, vunitcompile, vunitelaborate, vunitruntests, vunitgetreport, vunitgettestlog, vunitgettestwaveform, vunitexportjson, vunittestdependencies). Use when the user asks to run or compile VUnit tests, find…
Structured code review for bugs, regressions, tests, and standards. Report-only by default; apply fixes only when the current user or upstream caller explicitly requests review-and-fix. mode:agent is always report-only.
Produce a coverage-checked report of the entire Memex conversation corpus when the user asks to analyze, organize, or summarize all Codex history. Use deterministic totals first and label unfinished backfill honestly.
Run the full verification suite required before finishing any change to the thechrisgrey codebase. Covers lint, typecheck, tests, formatting, and AGENTS.md validation.
Runs an end-to-end test from a Jira ticket number. Reads the ticket, executes the browser test steps it contains with Playwright, optionally scores any AI or text output with the RetriEval MCP, and writes a pass/fail result back to the ticket. Use this WHENEVER the user gives a ticket key and asks to test, verify, QA…
Run automated mobile app tests on BrowserStack App Automate using runAppTestsOnBrowserStack and takeAppScreenshot MCP tools. Use for Appium/XCUITest/Espresso testing.
★not rated 0 2mo agoA44 tokens
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: