Benchgen
Models/openai/

GPT-5 Codex

DraftPublic

Model Details

GPT-5 Codex

Organization License Released Focus

Quick answer: GPT-5 Codex is OpenAI's September 2025 coding-specialised model, scoring 74.5% on SWE-Bench Verified. It represents OpenAI's revival of the "Codex" brand for code-focused models within the GPT-5 generation, targeting real-world software engineering tasks.

At a Glance

Where GPT-5 Codex leads

  • 74.5% SWE-Bench Verified — strong real-world software engineering
  • Specialised for code generation and software engineering tasks
  • GPT-5 generation quality in a code-focused variant

Where it lags

  • September 2024 knowledge cutoff — earlier than GPT-5.1 Codex
  • Superseded by GPT-5.1 Codex (February 2026) for most coding tasks
  • Proprietary: no open weights
  • Text-only (coding focus)

Best for: Software engineering agentic tasks using the Codex interface; production coding pipelines; automated code review and PR fixing.

What GPT-5 Codex Is

GPT-5 Codex (released September 2025) is OpenAI's code-specialised model in the GPT-5 generation — a revival of the Codex brand last seen in the original Codex/Copilot era. It is distinct from the general-purpose GPT-5 models, with training and fine-tuning specifically targeting software engineering.

The 74.5% SWE-Bench Verified score places it competitively with DeepSeek-V3.2 Speciale (73.1%) and just above the threshold that makes models practical for real-world repository-level coding tasks. GPT-5.1 Codex (February 2026) provides a modest improvement to 73.7% — suggesting the generation gap between Codex versions was relatively small.

Specifications

FieldValue
OrganizationOpenAI
LicenseProprietary (API only)
Release dateSeptember 2025
Knowledge cutoffSeptember 2024
ModalityText / Code
FocusSoftware engineering (SWE-Bench)

Pricing

Available via OpenAI API — refer to the OpenAI pricing page for current Codex model rates.

Public Benchmark Scores

BenchmarkScoreSourceDate
SWE-Bench Verified74.5%Benchgen evaluation2025-09

GPT-5 Codex vs Alternatives

ModelSWE-Bench VerifiedLicense
GPT-5 Codex74.5%Proprietary
GPT-5.1 Codex73.7%Proprietary
DeepSeek-V3.2 Speciale73.1%MIT
DeepSeek-V4 Flash Max79%Proprietary

GPT-5 Codex vs DeepSeek-V3.2 Speciale: +1.4pp SWE-Bench, proprietary vs MIT. For open-weight SWE, V3.2 Speciale is preferred. For maximum SWE-Bench: DeepSeek-V4 Flash Max (79%).

Frequently Asked Questions

What is GPT-5 Codex? OpenAI's September 2025 coding-specialised model scoring 74.5% SWE-Bench Verified. A GPT-5 generation revival of the Codex brand for software engineering tasks.
What is the difference between GPT-5 Codex and GPT-5.1 Codex? GPT-5.1 Codex (February 2026) is the updated version — slightly lower SWE-Bench (73.7% vs 74.5%) but with a more recent September 2025 knowledge cutoff.

Specs from OpenAI's GPT-5 Codex release (September 2025) and Benchgen evaluations. Last updated 2026-07-24.

Benchmark Leaderboards

This model isn’t on any benchmark leaderboard yet.