Independent register of AI agents & harnesses · no sponsored placementsEdition 2026-09 · 107 entries · evidence to 26 Sept 2026
AgentsWisdom

Compare / Claude Code vs Codex (CLI and app)

Claude Code vs Codex (CLI and app)

Claude Code rates 73 against Codex (CLI and app)'s 71 in this edition. The table lines up every dimension we record; the leading value in each numeric row is marked. Add a third or fourth entry.

Dimension
Claude CodeAnthropic
Rating
AgentsWisdom score7371
Confidencehighhigh
Rank in categoryNo. 2 of 16No. 5 of 16
Pillars
Adoption & momentum9792
Experts (independent reviews)6855
Crowd (community sample)4667
Community posts sampled1715
Experts vs crowdCrowd cooler by 22Crowd warmer by 12
Adoption
GitHub stars148k127k
npm weekly downloads12M19M
PyPI monthly downloadsn/en/e
VS Code installs26M15M
Reported usersn/e5M
Reported customer organisationsn/en/e
Releases, last 90 days7535
Benchmarks (shown, not scored)
Terminal-Bench 2.0No. 31 of 43 · 57.98%
Claude Code + Claude Opus 4.6
No. 4 of 43 · 82.25%
Codex CLI + GPT-5.5
SWE-benchnot listednot listed
Independent reviews88
Trust & safety
Overall gradeBB
Permission modelCA
Data access scopeAA
Data storageBB
Data retention & trainingCC
Incident historyCC
ComplianceAA
TransparencyBA
Price
PriceFrom $20Free, paid from $8
Plans
  • Free: $0 (see notes)
  • Pro: $20 / monthly
  • Max 5x: $100 / monthly
  • Max 20x: Custom
  • Team Standard seat: $25 / monthly
  • Team Premium seat: $125 / monthly
  • Enterprise / API: Custom
  • Free: Free
  • Go: $8 / monthly
  • Plus: $20 / monthly
  • Pro: $100 / monthly
  • Business: $20 / monthly
  • Enterprise & Edu: Custom
  • API key: Custom
Product
CategoryCoding agentsCoding agents
MakerAnthropicOpenAI
LicenceproprietaryApache-2.0
Use it viaCLI, IDE extension, Desktop app, Web app, Mobile app, APICLI, Desktop app, IDE extension, Web app, API
Runs onmacos, linux, windows, webmacos, linux, windows, web
Models it drivesClaudeGPT
Key features
  • Terminal CLI plus VS Code and JetBrains extensions, desktop, web and mobile surfaces
  • Permission modes: Manual, Accept Edits, auto mode (classifier-reviewed) and others
  • Opt-in OS-level Bash sandbox (Seatbelt on macOS; Linux and WSL2 supported; native Windows not)
  • Cloud sessions in isolated Anthropic-managed VMs
  • Subagents, hooks, skills, plugins and MCP server support
  • Claude Agent SDK for building custom agents on the same harness
  • Open-source Rust CLI (Apache-2.0) installable via npm, Homebrew or standalone installers
  • Desktop app with in-app diff review and inline comments
  • IDE extension for VS Code, Cursor and Windsurf
  • OS-level sandbox on by default locally, with configurable approval policy and network off by default
  • Optional auto-review mode that routes approval requests through a reviewer agent
  • Cloud tasks that run in OpenAI-hosted environments
Verdicts
Praised for
  • Widely treated by practitioners as a reference coding agent and daily driver (Simon Willison; HN comments)
  • Deep configurability through CLAUDE.md, skills and hooks rewards investment (Composio 100+ hour comparison)
  • Explains its own changes well enough to support review workflows (Every / Kieran Klaassen)
  • Generous usage within ChatGPT subscriptions compared with rivals, per many community comments (HN)
  • Strong Terminal-Bench 2.0 result: rank 4 of 43 harnesses
  • Safe-by-default local sandbox and approval model (docs)
  • Delegation and review workflow in the app praised by reviewers (Every; Composio)
Criticised for
  • Frequent complaints about usage limits and token burn on subscription plans (HN, r/ClaudeCode)
  • Long record of security advisories in the CLI (about 30 GHSA/CVE entries in 24 months) plus a March 2026 source-map leak
  • Auto mode default relies on a classifier that has been bypassed in published research (Johann Rehberger via Simon Willison)
  • Terminal UI performance complaints such as freezing and memory use (Douglas Mendes; HN)
  • Usage-limit and rate-limit changes cause sharp complaints (GitHub issue #28879 with 364 upvotes)
  • Reported bugs with serious side effects, e.g. $HOME deletion in full-access mode (Simon Willison, July 2026) and excessive SSD writes from logs
  • Weak front-end/UI output reported by some users (r/codex)
Evidence
Evidence collected26 Sept 202626 Sept 2026
Sources cited43