Skill

Delegate bounded coding tasks to Codex CLI subagent

Delegates bounded coding, review, or verification tasks to a separate, sandboxed Codex CLI session via codex exec.

Works with codexchatgptgit

91
Spark score
out of 100
Updated 15 days ago
Version 1.0.0

Add to Favorites

Why it matters

Offload self-contained coding, review, or verification tasks to an autonomous Codex CLI session that runs in parallel, streams progress, and delivers final results without polluting your main conversation context.

Outcomes

What it gets done

01

Launch non-interactive Codex sessions with clear success criteria and file ownership

02

Run multiple independent coding tasks in parallel using separate git worktrees

03

Collect and review Codex's changes and final deliverable message

04

Resume previous Codex sessions for follow-up instructions or refinements

Install

Add it to your toolbox

Run in your project directory:

curl -fsSL https://spark.entire.vc/get/ag-codex-subagent | bash

Overview

Codex CLI as a Subagent

Delegates a bounded coding, review, or verification task to a separate, sandboxed Codex CLI session (codex exec), which runs non-interactively and reuses the user's ChatGPT auth rather than an API key. Use it for self-contained tasks with clear success criteria, independent parallel work, or a second opinion on your own changes. Do not use it when the task needs conversation context that can't be written into the prompt.

What it does

Delegates a bounded coding, review, or verification task to a separate Codex CLI session - OpenAI's terminal coding agent. codex exec runs it non-interactively: it works autonomously in a sandbox, streams progress to stderr, and prints only the final message to stdout. Auth reuses the user's ChatGPT subscription, never an API key.

When to use - and when NOT to

Delegate a self-contained coding task with clear success criteria (fix, feature, refactor, review), several independent parallel tasks at once, or when you want a second opinion or independent verification of your own changes. Do NOT delegate tasks that need conversation context you can't fully write into the prompt - Codex sees nothing of the conversation, so the prompt must carry the goal, relevant paths, constraints, and how to verify completion.

Inputs and outputs

Preflight checks codex --version (install via npm i -g @openai/codex or brew install --cask codex if missing) and codex login status (exit 0 + "Logged in using ChatGPT" means ready; otherwise stop and tell the user to run codex login, a one-time browser OAuth - never read, print, or copy ~/.codex/auth.json). Launch:

OUT=$(mktemp /tmp/codex-out.XXXXXX)
codex exec \
  --cd /path/to/repo \
  --sandbox workspace-write \
  --output-last-message "$OUT" \
  "Full task prompt: goal, constraints, files to touch, definition of done." \
  </dev/null

</dev/null is mandatory when stdin isn't a real terminal, since Codex treats open stdin as extra context and waits forever for EOF. A long prompt can be piped via stdin instead, and the verbose stream can be backgrounded or wrapped in a subagent to stay out of the parent context; runs take minutes with no built-in timeout. Optional flags: -m <model> to override the model, --json for a JSONL event stream. Collect results with:

cat "$OUT"                            # final message = the deliverable
git -C /path/to/repo status --short   # see what Codex actually changed

Follow-up in the same session (cwd-filtered) uses codex exec resume --last "follow-up instruction" </dev/null.

Integrations

Wraps OpenAI's Codex CLI (@openai/codex, installable via npm or Homebrew) and reuses the host's ChatGPT auth. Can be wired as a Cursor-native /codex subagent by adding ~/.cursor/agents/codex.md pointing at this skill. For parallel runs, only genuinely independent tasks should be parallelized with upfront file ownership so results merge cleanly - one git worktree per Codex run, never two in the same tree. Known failure modes: a hang with no output means stdin was left open (kill and relaunch with </dev/null); a non-zero codex login status means the user must run codex login (don't work around it); a ChatGPT plan rate limit should be reported to the user, never retried in a loop; a "Not a git repo" error needs --skip-git-repo-check or a git init; network is blocked inside the workspace-write sandbox by default unless the task needs installs or API calls, in which case pass -c sandbox_workspace_write.network_access=true. Never use --dangerously-bypass-approvals-and-sandbox.

Who it's for

Agents or developers who want to hand off a bounded coding, review, or verification task to an independent CLI agent, run several independent tasks in parallel, or get a second opinion on their own changes - one task per launch, splitting big jobs into multiple launches, and always reviewing Codex's diff before declaring the task done.

Source README

Codex CLI as a Subagent

When to Use

  • Use when a bounded coding, review, or verification task can run in a separate Codex CLI session.
  • Use when parallel work needs explicit file ownership and a clear definition of done.

Codex CLI is OpenAI's terminal coding agent. codex exec runs it non-interactively:
it works autonomously in a sandbox, streams progress to stderr, and prints only the
final message to stdout. Auth reuses the user's ChatGPT subscription - never an API key.

When to delegate

  • Self-contained coding task with clear success criteria (fix, feature, refactor, review).
  • Parallel work: several independent tasks at once (see Parallel runs).
  • Second opinion / independent verification of your own changes.

Do NOT delegate tasks that need conversation context you can't fully write into the prompt.

Preflight

codex --version       # missing? npm i -g @openai/codex  (or: brew install --cask codex)
codex login status    # exit 0 + "Logged in using ChatGPT" = ready

Not logged in → stop and tell the user to run codex login (one-time browser OAuth).
Never read, print, or copy credentials (~/.codex/auth.json).

Launch

OUT=$(mktemp /tmp/codex-out.XXXXXX)
codex exec \
  --cd /path/to/repo \
  --sandbox workspace-write \
  --output-last-message "$OUT" \
  "Full task prompt: goal, constraints, files to touch, definition of done." \
  </dev/null
  • </dev/null is MANDATORY when stdin is not a real terminal (background shells,
    scripts): codex treats open stdin as extra context and waits forever for EOF.
  • Codex sees NOTHING of your conversation. Put all context in the prompt:
    goal, relevant paths, constraints, and how to verify it's done.
  • Long prompt? Pipe it via stdin instead: codex exec [flags] - < /tmp/task.md.
  • Wrap the command in a background/Bash subagent if your host agent has one
    (Cursor: Task tool with a shell subagent) so Codex's verbose stream stays out
    of the parent context. Fallback: a plain background terminal.
  • Runs take minutes and have no built-in timeout - background it and monitor.
  • Optional: -m <model> to override the model, --json for JSONL event stream.

Collect results

cat "$OUT"                            # final message = the deliverable
git -C /path/to/repo status --short   # see what Codex actually changed

Follow-up in the same session (run from the same cwd - resume filters by cwd):

codex exec resume --last "follow-up instruction" </dev/null

Parallel runs

Parallelize only genuinely independent tasks, and assign file ownership upfront so
results merge cleanly. One git worktree per Codex run - never two in the same tree:

git worktree add /tmp/wt-taskA -b codex/task-a
codex exec --cd /tmp/wt-taskA --sandbox workspace-write -o /tmp/outA.md "task A" </dev/null

Failure modes

  • Hangs forever with no output → stdin was left open. Kill it, relaunch with </dev/null.
  • codex login status non-zero → the user must run codex login. Don't work around it.
  • ChatGPT plan rate limit hit → report to the user; never retry in a loop.
  • "Not a git repo" error → add --skip-git-repo-check, or init a repo first.
  • Network is blocked inside the workspace-write sandbox by default. If the task
    needs it (installs, API calls): -c sandbox_workspace_write.network_access=true.
  • NEVER use --dangerously-bypass-approvals-and-sandbox.

Rules

  • One task per launch. Split big jobs into multiple launches.
  • Review Codex's diff yourself before declaring the task done.

Cursor-native wrapper (optional)

For auto-routing and /codex invocation inside Cursor, add ~/.cursor/agents/codex.md -
a custom subagent whose description is "delegates coding tasks to Codex CLI" and whose
body points at this skill.

Limitations

  • Adapted from davidondrej/skills; verify local paths, tools, credentials, and agent features before acting.
  • For commands, remote access, scheduling, browser automation, or file-changing workflows, get explicit user approval and confirm the target environment first.

FAQ

Common questions

Discussion

Questions & comments · 0

Sign In Sign in to leave a comment.