analyzer
01d-o-hub/github-template-ai-agents
Agent Codex
Analyze benchmark results from the eval pipeline to surface actionable patterns for skill improvement.
A template repository for GitHub projects using AI agents (Claude, Copilot Chat, OpenAI Codex, ...) with automated label setup, branch protection, and workflow best practices.
d-o-hub/github-template-ai-agents
Agent Codex
Analyze benchmark results from the eval pipeline to surface actionable patterns for skill improvement.
d-o-hub/github-template-ai-agents
Agent Codex
Blind comparison of two skill versions to determine which produces higher quality outputs.
d-o-hub/github-template-ai-agents
Agent Codex
Grade eval assertion results against actual skill outputs. Expects structured input and returns deterministic pass/fail with concrete evidence.
d-o-hub/github-template-ai-agents
Agent Claude Code
Create new Claude Code agents with proper format, YAML frontmatter, system prompts, and tool configuration. Invoke when you need to build specialized sub-agents for autonomous task execution.