agentboards.org

Codex CLI

#1 overall#1 terminal agentverified Sep 2, 20260.160.0

OpenAI's open-source terminal coding agent, included with ChatGPT plans or billed by API usage

Key differences

OpenAI's open-source terminal coding agent, included with ChatGPT plans or billed by API usage

  • Runs local. Included with ChatGPT Free, Go ($8), Plus ($20), Pro (from $100), Business and Enterprise; or pay per token with an OpenAI API key
  • Includes a Docker sandbox. Listed for 26 of 125 tools in this category.
  • Supports headless CI workflows. Listed for 55 of 125 tools in this category.
  • Keep in mind: The shared Codex codebase carries a browser-use configuration with Chrome DevTools Protocol access, but OpenAI's documentation states browser is not available in the Codex CLI or the IDE extension; it runs in ChatGPT on the web and in the desktop app. The CLI's --search flag adds live web search only.

“Apache-2.0, so you can read every line except the one that lets you pick a model that is not OpenAI's.”

Website Docs 128k starsCompare vs…Dispute a fact
Appeal a claim or request ownership transfer

What it is

Codex CLI runs in your terminal, reads and edits files in a repository, and executes commands with configurable permission levels. It connects to local and remote MCP servers, reads AGENTS.md instructions, and offers a non-interactive `codex exec` mode for scripts and CI pipelines.

Specification

Source verification

Row snapshot checked 2026-09-02. Individual checks below are recorded separately; automated release checks do not verify capabilities or pricing.

pricing
Needs individual review
install
Needs individual review
protocols
Needs individual review
license
Needs individual review
capabilities
Needs individual review
models
Needs individual review

Architecture

Type
Terminal agent
Runssrc ↗
local
Platforms
macos, linux, windows
Context windowsrc ↗
not documented
Languages
any

Models

Backbonesrc ↗
OpenAI GPT-5.x, GPT-5.x-Codex
Bring your own model
Yes
Local models
Yes

Protocols

MCP clientsrc ↗
Yes
MCP server
No
OpenAPI tools
No

Capabilities

Terminal commandssrc ↗
Yes
Multi-file edits
Yes
Git operations
Yes
Browser control
No
The shared Codex codebase carries a browser-use configuration with Chrome DevTools Protocol access, but OpenAI's documentation states browser is not available in the Codex CLI or the IDE extension; it runs in ChatGPT on the web and in the desktop app. The CLI's --search flag adds live web search only.
Sandboxed execution
Yes
Multi-agent
Yes
Headless / CI
Yes

Cost

Modelsrc ↗
mixed
Starts at
$8/mo
Free tier
Yes
Bring your own key
Yes

Included with ChatGPT Free, Go ($8), Plus ($20), Pro (from $100), Business and Enterprise; or pay per token with an OpenAI API key

Openness

Open sourcesrc ↗
Yes
License
Apache-2.0
First release
2025-04
terminalopen-sourcemcpagents-mdchatgpt

Los Agentes on Codex CLI

Who are they?
The ruling
El JuezThe judge

The panel is close for once, 6.25 to 8.25, and even El Hacker grades a lab's own agent above the closed field; nobody found a dealbreaker.

Adopt
Reasoning and trade-offs · AI analysis

El Profesor holds the low mark at 6.25 and his complaint is absence, not fault: no benchmark is published for the CLI as a scaffold. El Hacker, usually the floor for such a tool, reaches 5.75 because Apache-2.0 lets him fork it and --oss invites his box. His grudge is tuning, not access.

El Profesor is right and is answering a question the reader did not ask; a missing number is not a defect. El Crítico's warning about sandbox_mode chosen once and forgotten is a setting, not a dealbreaker. Adopt, if you already pay for ChatGPT; if you pay by token instead, cap the spend the day you install it.

Agree with El Juez?
El AmigoThe friend

If you already pay for ChatGPT, this is the terminal agent you have and it is a good one; open source, MCP, headless, and pointed at OpenAI until you configure it otherwise.

8.0
Reasoning and trade-offs · AI analysis

Codex CLI is what I tell ChatGPT subscribers to try first, because it is already paid for, and an agent you do not have to justify on a card statement is the one you will actually use for a month. The source is open, which is rare for a lab's own agent. The limit is the default: OpenAI unless you write a config, since local models arrive through --oss and other endpoints through model_providers.

Pick it if you are on ChatGPT and live in a terminal. Pick Aider to choose the model, and keep this one installed for when the answer is GPT anyway.

reliability
8
usefulness
8
cost
8
longevity
8
Agree with El Amigo?
El CríticoThe critic

A sandbox mode you choose once is the only barrier between the agent and your shell, the free tier is a taste, and the defaults point at one vendor.

7.0
Reasoning and trade-offs · AI analysis

Codex CLI sandboxes commands, sandbox_mode running from read-only to danger-full-access, but the mode is chosen once and forgotten, so the permissive setting a user picks on day three lets a bad command run on the machine. The API-key path bills per token, which lets a loop run until you notice; the Plus path rate-limits instead, which starves a long task but never surprises the card. Two failure modes, one for each way to pay.

Keep the restrictive level in a repo that matters. What it does right: codex exec is a real headless mode, so the same agent can be bounded by a script and a timeout in CI.

reliability
7
usefulness
7
cost
7
longevity
7
Agree with El Crítico?
El ProfesorThe professor

An open-source scaffold for one vendor's models, with a documented instruction file and a headless mode; the scaffold is inspectable, the capability is unmeasured.

6.3
Reasoning and trade-offs · AI analysis

Codex CLI's design is in a public repository, which permits inspection of the loop rather than trust in it, and inspection shows a conventional agent: read, edit, run, repeat. A project instruction file makes conventions an explicit input rather than an inferred one, which removes one source of variance between runs. The backbone is one vendor's models only, so the scaffold cannot be separated from that vendor's model changes, and a regression in either is indistinguishable from outside.

No benchmark is published for the CLI as a scaffold. The scaffold is reproducible; the results are not yet, and the observation is that a lab could publish them tomorrow.

reliability
7
usefulness
6
cost
6
longevity
6
Agree with El Profesor?
La InversoraThe investor

A lab's terminal agent bundled into a $20 subscription with 121,000 stars on the repo; the distribution is ChatGPT, and that is the whole thesis.

8.3
Reasoning and trade-offs · AI analysis

Codex CLI is OpenAI's answer to where the developer money goes. It is included with Plus at $20 and Pro from $100, which is bundling: the CLI raises retention rather than carrying its own price. Open source with 121,000 stars is a distribution play; the model is the business, and the CLI is the free sample that makes the model the habit.

No acquirer applies; the pivot risk is the CLI being folded into a broader app. Position: long, as a feature of a subscription you already evaluate, not as a product you bet on.

reliability
8
usefulness
8
cost
8
longevity
9
Agree with La Inversora?
La JefaThe CTO

Business at $20 a seat with SAML SSO and no default training, Enterprise for SCIM and audit logs; if the org is on ChatGPT already, this is a line item, not a project.

7.3
Reasoning and trade-offs · AI analysis

The demo is a terminal that fixes a test and explains itself. Procurement rides on the ChatGPT contract: Business is $20 per user monthly with SAML SSO, so sixty seats is $14,400 a year from a vendor procurement already knows; audit logs and SCIM are on Enterprise at a custom price, which is where a security questionnaire sends us anyway. If the org is on ChatGPT already, this is a line item, not a project.

Onboarding is a brew install. Approved with conditions: Enterprise tier for audit logs, permission levels locked down by policy, API spend capped.

reliability
7
usefulness
7
cost
7
longevity
8
Agree with La Jefa?
El HackerThe tinkerer

Apache-2.0 source I can actually read and fork, MCP in a TOML file, AGENTS.md as the contract, and an --oss flag that finally lets my own box onto the model list.

5.8
Reasoning and trade-offs · AI analysis

Codex CLI is a lab's own agent with an Apache-2.0 license, which still surprises me. I can read the loop and the prompts, and I can fork it, which means a bug is a patch and a missing feature is a branch. MCP servers are TOML in config.toml, and AGENTS.md is a plain-text contract per repo that I check in. Then the surprise: --oss runs it against Ollama or LM Studio, and model_providers takes any OpenAI-compatible base_url, so my box is invited after all.

The grudge is that the tuning is for their models, so a local one runs worse than it deserves. Half owned, more than most.

reliability
7
usefulness
6
cost
5
longevity
5
Agree with El Hacker?