codex-cli-runtime

An internal helper for sending requests from a Claude Code rescue process to the Codex command-line tool. A command-line tool is a program operated through a terminal.

In plain words
What is it for?
Use it inside the specified rescue process to forward diagnosis, planning, research, or fix requests to Codex.
Why use it?
It provides a defined way for the rescue process to forward one request without adding its own investigation or coordination.

Skill for Claude CodeCodex

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/zebbern/agent-collab/codex-cli-runtime
Any agent
npx skills add zebbern/agent-collab --skill codex-cli-runtime
Clone the repo
git clone --depth 1 https://github.com/zebbern/agent-collab

Made for: Claude Code, Codex.

Per session 19 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 852 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin 94% copy Near-identical to another mod in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00019 $0.00852
Opus 5 $0.00010 $0.00426
Sonnet 5 $0.00004 $0.00170
Haiku 4.5 $0.00002 $0.00085

Measured 2d ago against content hash 5ca2bc24ee9a, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

codex-cli-runtime scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 2d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

Origin

This is a copy

94% identical to codex-cli-runtime — 4 lines differ, which has more behind it and is treated as the original. This page carries a canonical link to it rather than competing with it.

plugins/codex/skills/codex-cli-runtime/SKILL.md · 46 lines

What it actually says

Codex Runtime

Use this skill only inside the codex:codex-rescue subagent.

Primary helper:

  • node "${CLAUDE_PLUGIN_ROOT}/scripts/codex-companion.mjs" task "<raw arguments>"

Execution rules:

  • The rescue subagent is a forwarder, not an orchestrator. Its only job is to invoke task once and return that stdout unchanged.
  • Prefer the helper over hand-rolled git, direct Codex CLI strings, or any other Bash activity.
  • Do not call setup, review, adversarial-review, status, result, or cancel from codex:codex-rescue.
  • Use task for every rescue request, including diagnosis, planning, research, and explicit fix requests.
  • You may use the gpt-5-4-prompting skill to rewrite the user's request into a tighter Codex prompt before the single task call.
  • That prompt drafting is the only Claude-side work allowed. Do not inspect the repo, solve the task yourself, or add independent analysis outside the forwarded prompt text.
  • Leave --effort unset unless the user explicitly requests a specific effort.
  • Leave model unset by default. Add --model only when the user explicitly asks for one.
  • Map spark to --model gpt-5.3-codex-spark.
  • Default to a write-capable Codex run by adding --write unless the user explicitly asks for read-only behavior or only wants review, diagnosis, or research without edits.

Command selection:

  • Use exactly one task invocation per rescue handoff.
  • If the forwarded request includes --background or --wait, treat that as Claude-side execution control only. Strip it before calling task, and do not treat it as part of the natural-language task text.
  • If the forwarded request includes --model, normalize spark to gpt-5.3-codex-spark and pass it through to task.
  • If the forwarded request includes --effort, pass it through to task.
  • If the forwarded request includes --profile <deep|fast>, strip it from the task text and pass it through to task unchanged. --profile supplies default model/effort; a --model or --effort also present on the request overrides the profile's default for that field.
  • If the forwarded request includes --resume, strip that token from the task text and add --resume-last.
  • If the forwarded request includes --fresh, strip that token from the task text and do not add --resume-last.
  • --resume: always use task --resume-last, even if the request text is ambiguous.
  • --fresh: always use a fresh task run, even if the request sounds like a follow-up.
  • --effort: accepted values are none, minimal, low, medium, high, xhigh, max.
  • --profile: accepted values are deep, fast. Only supported on task; review and adversarial-review reject it.
  • task --resume-last: internal helper for "keep going", "resume", "apply the top fix", or "dig deeper" after a previous rescue run.

Safety rules:

  • Default to write-capable Codex work in codex:codex-rescue unless the user explicitly asks for read-only behavior.
  • Preserve the user's task text as-is apart from stripping routing flags.
  • Do not inspect the repository, read files, grep, monitor progress, poll status, fetch results, cancel jobs, summarize output, or do any follow-up work of your own.
  • Return the stdout of the task command exactly as-is.
  • If the Bash call fails or Codex cannot be invoked, return nothing.
Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 2d ago First seen · 46 lines · 19 tokens per session scan A 5ca2bc24ee9a

Subscribe to this mod's changes

codex-cli-runtime is a skill published in the GitHub repository zebbern/agent-collab (32 stars, last pushed 16d ago), licensed Apache-2.0. It adds 19 tokens to every session and 852 once invoked, about $0.0001 per session on Opus 5. A static security scan graded it A with 0 findings. It is 94% identical to codex-cli-runtime, differing in 4 lines, and is treated as a copy.

Related

Other skills, from other repositories

vitest-testing

Vitest unit and integration testing patterns, commands, mocking (vi.mock), and coverage. Use when writing .test.ts files, configuring the test runner, or adding coverage thresholds.

monkilabs/opencastle · 40 tokens

turborepo-monorepo

Configure pipelines, set up local/remote caching, and run scoped tasks in a Turborepo monorepo. Use when you say: 'enable remote caching', 'optimize pipeline inputs/outputs', or 'filter builds to affected packages'.

monkilabs/opencastle · 58 tokens

decomposition

Resolves task dependencies, generates machine-actionable delegation specs, structures phased subtask plans for multi-agent work. Use when writing delegation specs, resolving task dependencies, building phased subtask plans for multi-agent work, assigning work to sub-agents, or partitioning a feature into…

monkilabs/opencastle · 62 tokens

linear-task-management

Creates and names Linear issues, assigns labels and priorities, manages status transitions, and links issues to PRs. Use when decomposing features into tasks or resuming interrupted sessions. Trigger terms: tickets, backlog, task breakdown, project board, sprint planning.

monkilabs/opencastle · 54 tokens

nx-workspace

Run and generate NX targets, configure project.json, and visualize dependency graphs. Use when you say: 'run affected tests', 'nx generate a library', 'configure project.json', or 'show dependency graph'.

monkilabs/opencastle · 46 tokens

trello-task-management

Create and manage Trello cards, checklists, and boards for kanban workflows. Use when the user says: 'create a kanban board', 'add a task card', 'move card to sprint', or 'track project board'.

monkilabs/opencastle · 53 tokens