analyzer
313harshitsinghbhandari/domain-expansion
Agent Codex
Analyze blind comparison results to understand WHY the winner won and generate improvement suggestions.
945 that install into Codex. Same mod, same page, whichever agent you run: a URL per tool would split one page into five that compete.
harshitsinghbhandari/domain-expansion
Agent Codex
Analyze blind comparison results to understand WHY the winner won and generate improvement suggestions.
harshitsinghbhandari/domain-expansion
Agent Codex
Compare two outputs WITHOUT knowing which skill produced them.
harshitsinghbhandari/domain-expansion
Agent Codex
Evaluate expectations against an execution transcript and outputs.
Agent Codex
Records one open question's embedded recommendation as its answer in the current milestone's requirements.md, leaving the edit uncommitted for the orchestrator to commit. Invoke with the target question's Short Title as the prompt. Dispatched per-question by the answer-all-open-questions-with-recommendation sweep; not…
Agent Codex
Completes a single named task from the current milestone's TASKSTODO.md, verifies success criteria, and updates the task list. Invoke with the task's.
Agent Codex
Read-only recommendation subagent for a single open question — the non-interactive twin of discuss-open-question. Given one question's Short Title plus context, it grounds in the live project read-only, produces the alternatives + a single recommendation, and returns them as the block's XML sub-elements (one per…
sagar-shirwalkar/servicenow-atlas
Agent Codex
Analyze blind comparison results to understand WHY the winner won and generate improvement suggestions.
sagar-shirwalkar/servicenow-atlas
Agent Codex
Compare two outputs WITHOUT knowing which skill produced them.
sagar-shirwalkar/servicenow-atlas
Agent Codex
Evaluate expectations against an execution transcript and outputs.
d-o-hub/github-template-ai-agents
Agent Codex
Analyze benchmark results from the eval pipeline to surface actionable patterns for skill improvement.
d-o-hub/github-template-ai-agents
Agent Codex
Blind comparison of two skill versions to determine which produces higher quality outputs.
Agent Codex
Analyze blind comparison results to understand WHY the winner won and generate improvement suggestions.
Agent Codex
Compare two outputs WITHOUT knowing which skill produced them.
Agent Codex
Evaluate expectations against an execution transcript and outputs.
Agent Codex
Analyze blind comparison results to understand WHY the winner won and generate improvement suggestions.
Agent Codex
Compare two outputs WITHOUT knowing which skill produced them.
Agent Codex
Evaluate expectations against an execution transcript and outputs.
Agent Codex
Analyze blind comparison results to understand WHY the winner won and generate improvement suggestions.
Agent Codex
Evaluate expectations against an execution transcript and outputs.
LucasMatuszewski/JSystems-SilkyCoders-1
Agent Codex
Analyze blind comparison results to understand WHY the winner won and generate improvement suggestions.
LucasMatuszewski/JSystems-SilkyCoders-1
Agent Codex
Compare two outputs WITHOUT knowing which skill produced them.
LucasMatuszewski/JSystems-SilkyCoders-1
Agent Codex
Evaluate expectations against an execution transcript and outputs.
Agent Codex
Analyze blind comparison results to understand WHY the winner won and generate improvement suggestions.
Agent Codex
Compare two outputs WITHOUT knowing which skill produced them.