grader
97Agent
Evaluate expectations against an execution transcript and outputs.
583 tagged python, measured the same way as everything else here.
Browse within: Multi-Agent 39plugin 39agent-framework 30game-development 28knowledge-graph 28litestar 28fastapi 25openclaw 21adversarial-testing 20clawdbot 20paper-audit 20research-tools 20scientific-writing 20markdown 19
Agent
Evaluate expectations against an execution transcript and outputs.
Agent Claude Code
Workflow execution and verification specialist. Runs validatebeforeexecute, executes, and verifies outputs. Tests only; never modifies.
Agent Claude Code
Asset acquisition specialist. Downloads, verifies, and registers models. Provisions assets only; never modifies workflows or executes.
Agent Claude Code
Reconnaissance specialist. Discovers ComfyUI environment — nodes, models, custom nodes, system stats. Read-only authority. Never mutates state.
Agent
Primary use: Load this file for immediate context about the project.
Agent
Quick reference files for AI agents working on OneTool.
Agent
MCP server with single run tool for LLM Python code execution.
Agent
Expert in Jupyter Notebook and JupyterLab for interactive computing, data analysis, machine learning experimentation, and reproducible research. Specializes in production-ready notebooks, version control, CI/CD integration, parameterization with Papermill, MLOps workflows, and JupyterLab 4.4+ modern features including…
Agent Claude Code
Use PROACTIVELY after changing Rust FSL syntax, lowering, semantics, CLI commands, public Kernel contracts, or corpus specs. Reports missing coupled code, tests, docs, skills, generated artifacts, and changelog updates. Read-only.
Agent Claude Code
Use PROACTIVELY after changing Rust core/runtime/verifier/solver/refinement semantics. Audits symbolic BMC versus the solver-independent Monitor/BFS, false-negative risk, dependency boundaries, and cross-implementation evidence. Read-only on source; may run focused tests.
Agent Claude Code
Use PROACTIVELY after adding or changing a .fsl spec under specs/ or examples/. Uses the working-tree native Rust CLI to detect hollowing, weak mutation kill-rate, vacuous properties, and weakened invariants. Read-only on specs; may run verifier commands.
Agent Claude Code
Use this agent when you need infrastructure management, deployment automation, or operational excellence. This agent specializes in DevOps practices, cloud operations, monitoring setup, and maintaining reliable production systems. Context: Unifying multiple build scripts user: "I need help with unifying multiple build…
Agent Claude Code
Use this agent when you need infrastructure management, deployment automation, or operational excellence. This agent specializes in DevOps practices, cloud operations, monitoring setup, and maintaining reliable production systems. Context: When you need to deploy or manage infrastructure. user: "I need to deploy my…
Agent Claude Code
Use this agent when you need specialized assistance with image optimization specialist using imagemagick for web performance, format conversion, and responsive image generation. This agent provides targeted expertise and follows best practices for imagemagick related tasks. Context: When user needs optimize.image…
Agent
Implements the smallest safe fix for Project 4 and verifies it.
Agent
Strict read-only checker for Project 4 fix candidates.
Agent
Production benchmark runners invoke each host agent through its real CLI boundary. Deterministic mock runners (mock-synthetic) are reserved for offline harness tests only.
Agent
Install HM-Arch and configure Hermes to use the native HM-Arch Memory Provider. Hermes integration uses native plugin registration. hm-arch install hermes creates the HM-Arch plugin bridge, updates $HERMESHOME/config.yaml, and initializes the SQLite database.
Agent
Install HM-Arch as OpenClaw's native memory provider. OpenClaw integration uses the @hm-arch/openclaw-plugin TypeScript memory plugin backed by a persistent Python sidecar that exposes HM-Arch recall, capture, forget, and consolidation without per-query process startup.
MarcusJellinghaus/mcp-tools-py
Agent Claude Code
Commits and pushes code changes with pre-approved git operations.
MarcusJellinghaus/mcp-tools-py
Agent Claude Code
Approves issue for workflow status transition.
MarcusJellinghaus/mcp-tools-py
Agent Claude Code
Updates GitHub issue with refined content from analysis.
charles-adedotun/notifications-mcp-server
Agent Claude Code
Specialized agent for server-side development with modern backend frameworks, APIs, databases, and cloud infrastructure. Expert in building secure, scalable, and performant backend systems across multiple languages and platforms with comprehensive testing and monitoring strategies.
charles-adedotun/notifications-mcp-server
Agent Claude Code
Specialized agent for React frontend development with shadcn/ui, Tailwind CSS v4, and modern best practices. Expert in implementing design systems, ensuring accessibility, and building performant user interfaces following the 4-font-size, 2-weight typography system, 8pt grid spacing, and 60/30/10 color distribution…