tasker
553Plugin Claude Code
Bundles 6 skills, 8 commands, 8 agents, 3 hooks · 483 tokens together
Spec-Driven Development: specifications compiled into executable, verifiable behavior.
18,225 mods in this category, of every kind an agent can take. Each one carries what it costs per session, what the scan found, and whether it is the original.
Plugin Claude Code
Bundles 6 skills, 8 commands, 8 agents, 3 hooks · 483 tokens together
Spec-Driven Development: specifications compiled into executable, verifiable behavior.
Skill Claude CodeCodex
Recalibrate dash-p's recognition profile when a new Claude Code version ships. Drives the new TUI through a diverse SCENARIO BATTERY (short/long input, long output, code, markdown, tables, unicode, tool use), cross-checks every result against the session JSONL (ground truth), then fixes the profile/recognizer and…
Skill Claude CodeCodex
Expert setup assistant for the Stably Playwright SDK. Use this skill when installing Stably SDK in a new project, migrating from @playwright/test, or configuring Stably reporter for CI/CD. Triggers on tasks like "setup stably", "install stably sdk", or "configure playwright with stably".
Skill Claude CodeCodex
Reference guide for @gleanwork/mcp-server-tester — the Playwright-based testing and evaluation framework for MCP servers. Covers import paths, all 11 matchers, transport config, eval datasets, reporter setup, CLI commands, auth patterns, and common anti-patterns. Use when working with MCP server tests or evals.
Agent Claude Code needs its repo
End-to-end testing specialist using Playwright. Use PROACTIVELY for generating, maintaining, and running E2E tests. Manages test journeys, quarantines flaky tests, uploads artifacts (screenshots, videos, traces), and ensures critical user flows work.
rishigundakaram/cadquery-mcp-server
MCP server Claude CodeCodexCursor +2
MCP server "cad-verification" as configured in rishigundakaram/cadquery-mcp-server. Launched with /Users/rishigundakaram/.pyenv/shims/uv --directory /Users/rishigundakaram/Deskto.
Skill Claude Code
Use this skill when writing, reviewing, or deciding where to place tests in the Next.js frontend package. Covers which project (unit / integration / storybook) to use for a given scenario, what NOT to duplicate across projects, and file placement conventions for this package.
Command Claude Code
SAM end-to-end composer - runs plan, then tdd for every story, then comprehensive docs. The one-shot PRD-to-working-code experience.
Skill Claude CodeCodex needs its repo
Operate the belt CLI to evaluate headless coding agents (Claude Code, Cursor, Codex, Gemini, and others) end to end. Use when the user asks to write or run eval scenarios, compare agents, score outputs with rules or LLM judges, register a new agent adapter, interpret reports or benchmark cards, or set up evals in CI.…
Plugin Claude Code
Bundles 5 skills, 7 agents · 472 tokens together
Evaluate prompts and skills through parallel execution comparison in git worktrees. Prompt optimization with BP-001009 patterns, skill creation/update with quality grading, and blind A/B evaluation.
Skill Claude Code
Drive CrowTelemetry's OTLP ingest end-to-end — boot the real receiver, POST OTLP JSON with curl, inspect the SQLite db.
Skill Claude CodeCodex
Verify Python packages, CLIs, MCP servers, and Agent plugins through the exact installed artifact and real runtime contract. Use when source tests pass but a published release may omit modules, expose broken entry points, depend on local paths, fail without optional credentials, or behave differently after public…
Skill Claude CodeCodex needs its repo
Interaction with the iOS simulator using iosef, a CLI optimized for agent usage. Use when building or testing changes on the iOS Simulator — viewing the screen, tapping buttons, reading accessibility trees, finding elements by selector, asserting UI state, scripting multi-step test flows, installing and launching…
Skill Claude Code
Diagnose and test Claude Code skills against Anthropic's 7 principles. Scans SKILL.md files, checks 8 rules (gotchas, description, allowed-tools, file-size, structure, frontmatter, conflicts, usage-hooks), classifies skill types, generates prescriptions, and runs eval tests. Use when checking skill quality, auditing…
Skill Claude CodeCodex
Run visual regression tests, review screenshot diffs, and manage baselines on a Lastest instance via the @lastest/mcp-server MCP tools.
Skill Claude Code
Score way matching (embedding and BM25), analyze vocabulary, and validate frontmatter. Use when testing how well a way matches prompts, checking cosine similarity or BM25 scores, inspecting the embedding engine status, or validating way files.
Skill Claude CodeCodex
Operate Argos visual testing from the terminal with the argos CLI — inspect builds and snapshot diffs, submit reviews, request reviewers, post comments, inspect a test's flakiness and its recurring changes, ignore flaky test changes, configure a project and its contributors, manage a team's members, invites and email…
Skill Claude CodeCodex
Optimize anything through autonomous experimentation, against an evaluator you cannot grade yourself with. Use whenever the user wants to improve a number by iterating - raise accuracy, cut val loss, make training or inference faster, reduce latency or cost, beat a baseline, tune a pipeline or prompt, squeeze a…
Plugin Claude Code
Bundles 27 skills, 6 hooks · 997 tokens together
3 commands + 25 auto-trigger skills + self-evolving agent harness. Less to remember, more automation.
ever-works/directory-web-template
Skill Claude CodeCodex
Use when writing Playwright tests, fixing flaky tests, debugging failures, implementing Page Object Model, configuring CI/CD, optimizing performance, mocking APIs, handling authentication or OAuth, testing accessibility (axe-core), file uploads/downloads, date/time mocking, WebSockets, geolocation, permissions…
Skill Claude CodeCodex
Build evaluation harnesses for agents and chatbots — golden sets, deterministic tool-selection checks, LLM-as-a-Judge, Bedrock RAG evaluation jobs, CI gates, and online drift monitoring. Use when an agent needs a quality gate before merge or a quality alarm in production.
Skill Claude Code
A local experiment for testing a goal command that keeps an AI coding session working until a stated condition is judged complete. It uses a separate language model to evaluate the conversation after each assistant turn.
michaelpersonal/superpowers-cursor-rules
Cursor rule Cursor
Superpowers - A complete software development workflow with TDD, systematic debugging, and structured planning.
Cursor rule Cursor
When a soild fix - apply it without asking.
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: