Skip to content

OpenAI

GPT-5.3-Codex: price, context window, and caching

The February 2026 Codex-tuned coding model, which combined GPT-5.2-Codex's coding with stronger reasoning.

Released February 5, 2026 · Prices as of September 28, 2026

In OpenAI’s words

“The most capable agentic coding model to date.”

OpenAI docs: GPT-5.3-Codex

What OpenAI says it’s good at

  • GPT-5.2-Codex's coding with stronger reasoning and professional knowledge, running 25% faster for Codex users Source
  • Collaborating better while the agent is working Source

Facts

Specs and prices

FactGPT-5.3-Codex
MakerOpenAI
API model idgpt-5.3-codex
ReleasedFebruary 5, 2026
StatusPrevious generation
Context window400K tokens
Max output128K tokens
Open weightsNo
Input, per 1M tokens$1.75
Cache hit, per 1M$0.175
Cache write, per 1M$1.75 (same as input)
Output, per 1M tokens$14
Runs inCodex, Cursor, OpenCode, OpenRouter, and GitHub Copilot

Standard API rates in US dollars, as published by the maker on September 28, 2026. Batch and priority tiers, taxes, and subscription plans are not included.

Good to know

  • No longer selectable in Codex with ChatGPT sign-in since May 26, 2026. API-key use is unaffected.
  • It accepts up to 272K input tokens of its 400K context window.

Cost

What typical work costs

Example token counts at GPT-5.3-Codex’s published rates. On the agentic session, caching saves $3.15 against billing every token as ordinary input.

Example workload costs for GPT-5.3-Codex
Agentic coding session, 100K input, 400K written to cache, 2M read from cache, 50K output$1.93
Large one-off review, 150K input with no cache hits, 10K output$0.40
Output-heavy generation, 30K input, 80K output$1.17
A month of sessions, 110 sessions: 5 a day, 22 working days$211.75

Prompt caching

How OpenAI bills cached tokens

OpenAI

Prompt caching is on by default. From GPT-5.6 on you can also mark up to four explicit cache breakpoints, while GPT-5.5 and earlier cache automatically only.

From GPT-5.6 on, a cache write costs 1.25x the uncached input price and a cache hit costs 0.1x. GPT-5.5 and earlier add no charge for writing the cache: written tokens are billed as ordinary input, and a hit costs 0.1x on the models compared here.

On GPT-5.6 and later, a cached prefix stays reusable for at least 30 minutes after its last use, and caching starts at 1,024 input tokens.

Source: OpenAI: Prompt caching

GPT-5.3-Codex compared

  • Claude Sonnet 4.6 vs GPT-5.3-Codex

    Claude Sonnet 4.6 and GPT-5.3-Codex launched 12 days apart in February 2026. One takes 1M tokens of context, the other costs 46% less per session. Tradeoffs.

  • GPT-5.3-Codex vs Gemini 3.1 Pro Preview

    GPT-5.3-Codex and Gemini 3.1 Pro Preview cost within 4% of each other on a cached coding session. Context size, output limits, and access set them apart.

  • GPT-5.3-Codex vs GPT-5.2-Codex

    GPT-5.3-Codex and GPT-5.2-Codex cost the same to the cent. What differs is what OpenAI claims for each and where you can still run them in 2026.

  • GPT-5.4 vs GPT-5.3-Codex

    GPT-5.3-Codex is cheaper per token but takes 272K input tokens at most. GPT-5.4 reaches 1.05M, at higher rates past 272K. Which one fits your codebase.

Your own numbers

See what GPT-5.3-Codex really costs you.

everyaitoken reads your Codex, Cursor, OpenCode, and OpenRouter history on your Mac and prices every request at API rates, with what caching saved or cost. $9 once.

Launching soonSee the cache math