The Rundown AI homepage

Independent tool overview

Codex Max at a glance

Codex Max now means a quality-first reasoning setting that gives the selected current model more time on one difficult task. It is no longer the name of OpenAI's frontier coding model: the original GPT-5.1-Codex-Max model is deprecated and has been succeeded by newer Codex and GPT-5.6 models.

Visit the official Codex Max site ↗
Codex Max product preview
Current meaning
Quality-first reasoning mode for one task
Best for
The hardest single-agent coding or analysis work
Old model
GPT-5.1-Codex-Max is deprecated
Current flagship family
GPT-5.6 Sol, Terra, and Luna
API equivalent
reasoning.effort set to max
Separate price
None; plan limits or selected-model API rates apply

Overview

What Codex Max is

The name Codex Max has changed meaning. OpenAI launched GPT-5.1-Codex-Max as a standalone coding model for long-running agent tasks, but the current API catalog marks that model deprecated. Its old model ID and pricing should not be treated as OpenAI's present flagship coding offer.

In current Codex documentation, Max is a mode that gives whichever supported model you selected more time to reason about one task. OpenAI recommends it for the hardest problems when depth matters more than speed or usage. It differs from Ultra, which coordinates subagents in parallel rather than increasing one run's reasoning budget.

For API developers, the closest current control is max reasoning effort on GPT-5.6. It is a request setting—not a separate Codex Max model slug—and is billed at the selected model's token rates. The practical migration is to choose a current GPT-5.6 model and then test Max against Extra High on representative hard tasks.

Use cases

Who Codex Max is best for

The strongest fit depends on the job you need the product to complete, not the size of its feature list.

Difficult single-agent changes

Use Max for a tightly scoped but unusually hard implementation, debugging, optimization, review, or reasoning task.

Quality-sensitive work

Choose it when a marginal reliability gain matters more than latency and additional usage.

Measured escalation

Move from the default reasoning level to Max only after a representative task shows that lower efforts miss important requirements.

Capabilities

Core Codex Max features

1

More reasoning time

Max gives the selected model additional room to explore, plan, verify, and revise one difficult task.

2

Works with current model selection

The mode applies to a supported model you choose; it is not a separate current model named Codex Max.

3

Single-agent focus

Max deepens one run, unlike Ultra mode, which divides independent work among subagents.

4

API reasoning control

GPT-5.6 API requests can set reasoning effort to max while keeping the chosen Sol, Terra, or Luna model.

5

Configurable visibility

OpenAI notes that users who do not see Max may need to enable it in application settings.

Process

How the Codex Max workflow works

  1. Step 1

    Start with the right current model

    Choose Sol for the most difficult open-ended work, Terra for everyday production work, or Luna for clear high-volume tasks.

  2. Step 2

    Establish a baseline

    Run a representative task at the default or Extra High setting and define what success, latency, and usage look like.

  3. Step 3

    Escalate to Max

    Use Max only when the task remains difficult enough that additional reasoning could improve a valuable outcome.

  4. Step 4

    Compare the result

    Review correctness, completeness, tests, evidence, latency, and usage; keep Max only where it produces a measurable gain.

Cost

Codex Max pricing and free plan

Current Codex Max has no standalone price because it is a reasoning mode, not a separate product or model. ChatGPT-plan usage counts against the plan allowance and deeper reasoning can consume more of it. API usage is billed at the selected GPT-5.6 model's normal token rates, with Max potentially generating more reasoning tokens.

ChatGPT Plus

$20 per month

The standard individual Codex plan with GPT-5.6 access and extensible credits.

  • Codex across web, CLI, IDE, and supported apps
  • Usage is shared across local messages and cloud chats
  • Max can consume allowance faster than lower reasoning settings

ChatGPT Pro

From $100 per month

The higher-usage individual option with 5x or 20x the Plus Codex rate limits.

  • Pro 5x starts at $100 per month
  • Pro 20x is $200 per month
  • Higher plan limits do not make every Max run inexpensive or unlimited

GPT-5.6 Sol API

$4 input / $0.40 cached input / $20 output per 1M tokens

The flagship GPT-5.6 API model with max reasoning effort available as a request setting.

  • Max uses the same listed token rates as other reasoning efforts
  • More reasoning can increase output-token usage and latency
  • Prompts over 272K input tokens use higher long-context rates
  • Current promotional rate is stated to last at least through November 21, 2026

Legacy GPT-5.1-Codex-Max API

$1.25 input / $0.125 cached input / $10 output per 1M tokens

Historical pricing for the deprecated standalone model; not the recommended choice for new integrations.

  • Deprecated model ID: gpt-5.1-codex-max
  • Responses API only
  • 400K context window and 128K maximum output
  • Migrate to a current model before removal

Pricing checked . Check current pricing at the source ↗

Assessment

Codex Max strengths and limitations

Where it stands out

  • Provides an explicit quality-first escalation path for unusually difficult work.
  • Keeps model choice separate from reasoning depth.
  • Useful for hard tasks that do not divide naturally into independent subagent work.
  • API developers can benchmark max against lower efforts without changing model slugs.
  • OpenAI's current guidance clearly distinguishes Max from Ultra.

What to consider

  • The Codex Max name is ambiguous because it previously referred to the deprecated GPT-5.1-Codex-Max model.
  • Max increases latency and can consume substantially more plan allowance or API output tokens.
  • It is not guaranteed to improve every task, and routine work usually does not need it.
  • Max focuses one agent; tasks that split cleanly may finish faster with Ultra or another parallel workflow.
  • Availability can depend on the current Codex interface and settings.
  • The old gpt-5.1-codex-max model should not be selected for a new integration.
  • Even a higher-reasoning run can make incorrect changes, misuse tools, or miss requirements, so tests and review remain mandatory.

Compare

Codex Max alternatives

The right alternative depends on the specific output, workflow, controls and budget your project requires.

Coding

Codex

Use the main Codex page for the full product, interfaces, agent workflow, and current model choices.

Explore Codex

Coding

GPT-5.3-Codex-Spark

Choose Spark for the opposite tradeoff: near-instant, narrow coding iteration rather than maximum reasoning depth.

Explore GPT-5.3-Codex-Spark

Coding

Composer 2.5

Consider Cursor Composer 2.5 for cost-published long-running agent work inside the Cursor ecosystem.

Explore Composer 2.5

Consumer

Claude Sonnet 5

Compare Anthropic's current coding model when provider diversity and a different agent ecosystem matter.

Explore Claude Sonnet 5

Questions

Codex Max FAQs

Is Codex Max still available?

The current Max reasoning mode is active, but the old standalone gpt-5.1-codex-max API model is deprecated. These are different things.

What does Codex Max do now?

It gives the selected supported model more time to reason about one hard task. OpenAI recommends it when depth matters more than speed or usage.

Is Codex Max a model?

Not in the current Codex interface. Max is now a mode. GPT-5.1-Codex-Max was a separate model, but OpenAI's current catalog marks it deprecated.

What is the difference between Max and Ultra?

Max deepens one model run on a single task. Ultra coordinates subagents to work on separate parts of a task in parallel.

How much does Codex Max cost?

There is no separate Max fee. On ChatGPT plans it consumes the included Codex allowance, often faster than lower reasoning settings. In the API, the chosen model's normal token rates apply.

What should replace GPT-5.1-Codex-Max?

For current Codex use, select a GPT-5.6 model and enable Max only when needed. For API use, benchmark GPT-5.6 with reasoning effort set to max against Extra High and lower settings.

Bottom line

Our Codex Max verdict

Codex Max is useful, but only after correcting the name: it is now a depth setting, not OpenAI's latest coding model. Use a current GPT-5.6 model, establish a lower-effort baseline, and reserve Max for hard, high-value tasks where better reasoning is worth slower responses and greater usage. Migrate any code that still calls gpt-5.1-codex-max.

Visit Codex Max website ↗
The Rundown University

AI training for the future of work.

Get access to all our AI courses, hundreds of real-world AI use cases, live expert-led workshops, an exclusive network of AI early adopters, and more.

AI Courses

Get unlimited access to all of our current & upcoming industry-specific AI courses for the duration of your subscription.

Daily Guides

To keep up with the rapid pace of AI, our team publishes AI implementation guides daily. Our library contains 300+ practical use cases to automate real-world work.

Workshops

Join weekly, live, interactive sessions with industry leaders who are at the forefront of AI for hands-on implementation guidance and exclusive insights.

Community

Network with an exclusive community of AI-first professionals who are working smarter with AI. Learn how early adopters are using AI in their work and businesses.