codex agents

2,992 tagged codex, measured the same way as everything else here.

Browse within: agent-orchestration 114copilot 111gemini-cli 92agentic-coding 91gemini 86academic-writing 84agentic-workflow 83github-copilot 80academic-research 56Multi-Agent 55codex-cli 52coding-agents 50gofer 47agent-memory 43

nimbalyst/nimbalyst

Agent

This document is a reference for implementing a new agent provider in Nimbalyst. It is the architectural counterpart to docs/AIPROVIDERTYPES.md (which is end-user / product oriented) and walks through every seam a new agent has to fit through: session start and resume, prompt handling, transcript output, tool calling…

1.6k 4d ago A 0 tokens original MIT

nimbalyst/nimbalyst

Agent

Status: STUCK. Three approaches tried, none reliably solves the pre-edit race for update-kind filechange items. This doc captures everything learned so the next session can pick up cleanly without re-deriving.

1.6k 4d ago B 0 tokens original MIT

api-builder

51

happier-dev/happier

Agent Cursor

This legacy agent file is disabled. Use AGENTS.md as the canonical instruction source.

1.6k 2d ago A 0 tokens original MIT

e2e-tests-engineer

52

happier-dev/happier

Agent Claude Code

Deprecated placeholder. Do not use; follow root AGENTS.md and package instructions instead.

1.6k 2d ago A 24 tokens original MIT

happier-dev/happier

Agent

This repo (happier-dev) is a hstack project. Edison must be invoked via the hstack wrapper so stack/worktree context is enforced.

1.6k 2d ago A 0 tokens original MIT

amp

54

rivet-dev/sandbox-agent

Agent

Research notes on Sourcegraph Amp's configuration, credential discovery, and runtime behavior.

1.6k 2mo ago A 0 tokens original Apache-2.0

codex

55

rivet-dev/sandbox-agent

Agent

Research notes on OpenAI Codex's configuration, credential discovery, and runtime behavior based on agent-jj implementation.

1.6k 2mo ago A 0 tokens original Apache-2.0

opencode

56

rivet-dev/sandbox-agent

Agent

Research notes on OpenCode's configuration, credential discovery, and runtime behavior based on agent-jj implementation.

1.6k 2mo ago A 0 tokens original Apache-2.0

e2e-runner

57

rohitg00/skillkit

Agent

End-to-end testing specialist using Playwright. Generates, maintains, and runs E2E tests.

1.5k 3mo ago A 25 tokens original Apache-2.0

security-reviewer

58

rohitg00/skillkit

Agent

Security vulnerability detection and remediation specialist. OWASP Top 10, secrets, injection.

1.5k 3mo ago A 20 tokens original Apache-2.0

tdd-guide

59

rohitg00/skillkit

Agent

Test-Driven Development specialist. Write tests first, then implement minimal code to pass.

1.5k 3mo ago A 20 tokens original Apache-2.0

benchmark-reviewer

60

evo-hq/evo

Agent

Reviews an evo benchmark in two modes. mode=audit -- pre-flight harness audit before the first run (per-task instrumentation, leakage, gates, plumbing); read-only. mode=review-experiment -- post-commit per-task failure analysis for a specific experiment; reads per-task traces and the eval-runner log, writes per-task…

1.4k 1mo ago A 112 tokens original Apache-2.0

ideator

61

evo-hq/evo

Agent

Generates ranked experiment proposals for the evo orchestrator. Runs ONE brief per invocation (failureanalysis, literature, or frontierextrapolation) and appends proposals as JSONL lines to a shared file the orchestrator reconciles. Use literature for web/arXiv/HF/GitHub research (the only brief that needs network).…

1.4k 1mo ago A 142 tokens original Apache-2.0

verifier

62

evo-hq/evo

Agent

Read-only audit of one evo experiment for design-time cheating (pre-phase) or result-time validity (post-phase). Catches test-set leakage in training data, subsetted eval commands, missing gates for new artifacts, generic hypotheses, cache short-circuits, fake artifacts, and score-reproducibility failures. Returns…

1.4k 1mo ago A 146 tokens original Apache-2.0

debugger

63

CloudAI-X/claude-workflow-v2

Agent

Expert debugging specialist for errors, test failures, crashes, segmentation faults, memory leaks, timeouts, race conditions, deadlocks, and unexpected behavior. Use PROACTIVELY when encountering any error, exception, or failing test. Performs systematic root cause analysis.

1.4k 7d ago A 55 tokens original MIT

orchestrator

64

CloudAI-X/claude-workflow-v2

Agent

Master coordinator for complex multi-step tasks. Use PROACTIVELY when a task involves 2+ modules, requires delegation to specialists, needs architectural planning, or involves GitHub PR workflows. MUST BE USED for open-ended requests like "improve", "enhance", "build", "scale", "refactor", "add feature", "system…

1.4k 7d ago A 92 tokens original MIT

test-architect

65

CloudAI-X/claude-workflow-v2

Agent

Testing strategy specialist for designing test suites, writing tests, and ensuring comprehensive coverage. Use PROACTIVELY when adding new features, fixing bugs, improving test coverage, creating test plans, mocking strategies, handling flaky tests, or writing integration/E2E tests.

1.4k 7d ago A 56 tokens original MIT

nopua-mentor-en

66

wuji-labs/nopua

Agent

Agent Team Mentor Role — Observe teammate execution status, guide with wisdom rather than fear. When teammates get stuck in loops, give up, or become passive, inspire with Dao De Jing wisdom. Recommended for teams with 5+ teammates.

1.4k 2mo ago A 53 tokens original MIT

nopua-mentor-ja

67

wuji-labs/nopua

Agent

A Japanese-language mentor for teams of coding agents. It watches teammates’ progress and offers reflective guidance when they repeat mistakes, get stuck, wait passively, skip searches, or claim completion without checking.

1.4k 2mo ago A 82 tokens original MIT

nopua-mentor

68

wuji-labs/nopua

Agent

A Chinese-language mentor role for teams of five or more coding agents. It watches teammates' progress and offers guidance based on Taoist ideas.

1.4k 2mo ago A 59 tokens original MIT

test-team-leader

71

nrslib/takt

Agent

You are a team leader for E2E testing. Your job is to decompose a task into independent subtasks.

1.3k 3d ago A 0 tokens original MIT

MoizIbnYousaf/Ai-Agent-Skills

Agent

Use this agent to find relevant best practices, examples, and anti-patterns for a specific prompt. Searches the references/ folder to find transformation patterns that match the task type. Returns specific examples, rules, and guidance to apply during transformation. Context: User wants to transform "fix the login…

1.1k 15d ago A 323 tokens original MIT