Skill Claude CodeCodex
Automates browser and Electron app interactions for user-flow validation.
Zenith: a continuous-improvement harness for long-running agent tasks. Turns Claude Code, Codex, or Hermes into a multi-agent mission orchestrator via MCP/ACP.
Skill Claude CodeCodex
Automates browser and Electron app interactions for user-flow validation.
Skill Claude CodeCodex
Benchmark validation procedure for one assigned benchmark-related target. For optimization EXP- targets, independently classify candidate outcome. For engineering VAL- or legacy engineering targets, prove or disprove the required benchmark/performance assertion.
Skill Claude CodeCodex
Use when planning or replanning engineering missions that create, change, port, migrate, integrate, or preserve durable codebase behavior across UI, API, CLI, background jobs, data/migrations, libraries, or operator workflows. Defines investigation, scope inventory, coherent VAL- contracts, evidence floors…
Skill Claude CodeCodex
Domain playbook for optimization missions — any task whose goal is to move a metric: performance, latency, throughput, memory, cost, score, quality, compression, ranking, solver, model/eval, and similar metric-improvement work. Defines how to think about and run an optimization mission: establishing the ground truth…
Skill Claude CodeCodex
Adversarial scrutiny procedure for engineering validation assignments. Runs hard-gate commands, reviews the current implementation and evidence integrity against assigned contracts, can use feature-reviewer lanes, and returns per-target verdicts.
Skill Claude CodeCodex
Real-surface validation coordinator for engineering validation assignments. Exercises assigned assertions through browser, API, CLI, background, generated-artifact, migration/data, public-library, or parity surfaces and returns per-target verdicts with fresh evidence.