CodexHowSupport Us

GPT-5.3 Codex — Specs, Pricing & Limits

Specs

StatusPriced
Context window400,000 tokens
Max output128,000 tokens
Knowledge cutoffAug 31, 2025
Reasoning effort levelslow, medium, high, xhigh
Long-context ruleNot published

Pricing (short context)

Standard — input$1.75/MTok
Standard — cached input$0.175/MTok
Standard — cache writesNot published
Standard — output$14.00/MTok
Fast — input$3.50/MTok
Fast — output$28.00/MTok

Rate limits by usage tier

FreeNot supported
Tier 1500 RPM · 500,000 TPM · 1,500,000 batch queue
Tier 25,000 RPM · 1,000,000 TPM · 3,000,000 batch queue
Tier 35,000 RPM · 2,000,000 TPM · 100,000,000 batch queue
Tier 410,000 RPM · 4,000,000 TPM · 200,000,000 batch queue
Tier 515,000 RPM · 40,000,000 TPM · 15,000,000,000 batch queue

GPT-5.3 Codex is the model actually named for this site's subject — "optimized for agentic coding tasks in Codex or similar environments," in OpenAI's own words — and it's the one this whole site's most-read comparison exists to put next to the CLI's actual default, because the CLI doesn't default to it.

Its published context window is meaningfully smaller than the CLI default's, and its knowledge cutoff sits several months further back. What it offers in return: a per-token price well under half the default's on both input and output, and the only published reasoning-effort control on this site's entire roster — four distinct levels, stated on this model's own page and nowhere else. No other current model states any effort levels at all, published or otherwise.

It sits in the same rate-limit group as the frontier GPT-5.6 models, so choosing it over the CLI's default costs nothing in throughput — the tradeoff is specifically cost against context headroom and recency, not speed.

The long-context repricing rule that governs several other models isn't stated on this one's own page, and that absence is recorded honestly as not published rather than assumed to carry over from a sibling model. No Batch or Flex pricing is published for it either — Standard and Fast are the only two priced tiers this model offers.

The table below is computed directly from this site's facts module: pricing across every tier this model actually publishes, its four reasoning-effort levels, and its rate-limit table.

Verified 2026-08-09 against https://developers.openai.com/api/docs/models/gpt-5.3-codex.

Could not confirm: No Batch or Flex pricing, no cache-writes price and no >272K long-context repricing rule are published for this model. No Batch/Flex row appears in the pricing table at all for gpt-5.3-codex (checked the pricing page directly, not inferred from its absence on the model page).

Checked: https://developers.openai.com/api/docs/models/gpt-5.3-codex · https://developers.openai.com/api/docs/pricing