test
169LF-Decentralized-Trust-labs/gitmesh
Command Claude Code
Run the full test suite and summarize any failures.
4,592 tagged Testing, measured the same way as everything else here.
Browse within: agentic-workflow 42agentic-coding 33agent-orchestration 30browser-automation 30code-quality 28ai-development 27claude-plugin 25bug-bounty 24spec-driven-development 24ai-assistant 22agentic 21ai-workflow 21automated-testing 21gemini 21
LF-Decentralized-Trust-labs/gitmesh
Command Claude Code
Run the full test suite and summarize any failures.
Command Claude Code
A command for starting an IDE test instance by running the project’s Gradle task in the background. An IDE is a program developers use to write and test code.
Command Claude Code
Validate semantic proof graph EDN files against the Alethfeld schema.
Command
Start a feature: brainstorm a spec, get it approved, then drive it test-first through the pipeline. The one entry to implementation.
obra/superpowers-developing-for-claude-code
Command
Greet the user and demonstrate custom slash command functionality.
Command
Prove detected findings with runnable, defense-only evidence. For each finding the detect map carries, it writes a safe-path JavaScript test and a self-contained advisory under .lagune/proofs/. The test asserts secure behavior, never an exploit.
coleam00/dark-factory-experiment
Command
Pass-2 variant of dark-factory-synthesize-verdict. Reads -p2 node outputs (post-fix). Aggregates behavioral, security, code review, and static check results into an approve/requestchanges/reject verdict.
coleam00/dark-factory-experiment
Command
Final arbiter for Dark Factory PR validation. Aggregates behavioral, security, code review, and static check results into an approve/requestchanges/reject verdict.
coleam00/dark-factory-experiment
Command
Run Dark Factory validation — Python backend (ruff + mypy + pytest) + React frontend (tsc + biome + vitest).
dynamics365ninja/d365fo-mcp-server
Command Claude Code
Work the eval-loop corpus — rank failure clusters, fix the top actionable defect, validate against held-out, open a PR.
dynamics365ninja/d365fo-mcp-server
Command Claude Code
Run an eval case end-to-end on the VM (implement → build → score → record → roll back). VM/full-mode only.
Command Claude Code
Test all rust-docs-testing tools found in @rust-docs-mcp/src/service.rs.
Command
Design a warden acceptance plan through guided spec gathering. Detects existing planning artifacts (GSD, jira sprint, GitHub issue) and reads them as the spec source, or walks the user through adaptive Q&A. Self-bootstraps .warden/ on first invocation.
Command Claude Code
Review open pull requests for code quality, project alignment, and risks.
Command Claude Code
Analyze all open issues across GitHub Issues and Seeds, cross-reference with codebase health, and recommend the top 5 issues to tackle next.
Command Claude Code
Check a LaTeX coursework submission against the requirements in a supplied PDF assessment brief. Use when verifying format, required sections, word limits, or deliverables before submission. Not for general prose proofreading; use $proofread.
Command
Run multi-reviewer PRD stress test for build-readiness.
Command Claude Code
Fix errors in BCH E2E test (Pattern 1: P2PKH Single-sig).
Command Claude Code
Verify that the E2E scripts for the specified chain follow the documented transaction flow.
senaiverse/reactnative-expo-ai-agent-system-workflow
Command
Generates comprehensive test suite with ROI-based prioritization.
Command
You are the operator managing authentication credentials for the current engagement. The user's arguments specify the auth type and value.
Command
You are the operator resuming a previously interrupted engagement. The engagement directory and state files (scope.json, log.md, findings.md, cases.db) contain all the context needed to continue without repeating work.
Command
You are the operator providing a status summary of the current engagement.
Command
Command "tools_evasion" from SeaOf0/dsh-redteam-model, covering input, workflow, phase 1: tool understanding, phase 2: open source detection (if applicable) and phase 3: behavior analysis.