CRUCIBLE: a Claude-only adversarial expert workshop that stress-tests a scientific paper and then, opt-in, rebuilds it. ACT I (TRIBUNAL): from the paper PDF (plus a few public cited works it fetches), a topic-adapted fleet of expert referee subagents (every contested claim argued from at least two COMPETING…
Lets the agent drive a real browser: open pages, click, type, take screenshots and read the accessibility tree, using Playwright. Runs locally from the @playwright/mcp npm package.
Claude Code instructions for alexsds/ade-workflow, covering ade — agent-driven engineering, what this is, architecture, development and key principles.
Use this agent as a team member during /ade:execute to test and score features implemented by the generator. The evaluator uses pluggable rubrics and testing tools to grade work with hard thresholds. Context: The execute command is launching the agent team. user: "Build the approved plan" assistant: "Spawning the…
Use this agent as a team member during /ade:execute to implement features from an approved plan. The generator builds the app feature-by-feature, commits to git, and hands off to the evaluator for scoring. Context: The execute command is launching the agent team. user: "Build the approved plan" assistant: "Spawning…
This skill should be used when the user asks about evaluation methodology, scoring rubrics, testing tools, "how does scoring work", "evaluation criteria", "rubric format", "add a rubric", "create testing tool", "evaluate my feature", "run evaluation", "why did evaluation fail", or needs guidance on adversarial…
Use when asking about the ADE build process, how the generator works, iteration strategy, "why is the generator doing X", "how does building work", "commit conventions", "pivot vs refine", or understanding the implementation phase of the ADE workflow. This skill covers the Generator's methodology for implementing…
Use when the user wants to plan any work — building an app, adding a feature, fixing a bug, solving a problem, refactoring, or any task that benefits from thinking before doing. Triggers on "plan", "build me", "I want to", "fix this", "add a", "we need to", "how should we", or when the user describes something they…
1 5mo agoA115 tokens
originalMIT
At most 3 mods per repository are shown here — the rest are on their repository pages: