CodexHowSupport Us

GPT-5.4 Mini — Specs, Pricing & Limits

Specs

StatusPriced
Context window400,000 tokens
Max output128,000 tokens
Knowledge cutoffAug 31, 2025
Reasoning effort levelsNot published
Long-context ruleNot published

Pricing (short context)

Standard — input$0.750/MTok
Standard — cached input$0.075/MTok
Standard — cache writesNot published
Standard — output$4.50/MTok
Fast — input$1.50/MTok
Fast — output$9.00/MTok

Rate limits by usage tier

FreeNot supported
Tier 1500 RPM · 500,000 TPM · 5,000,000 batch queue
Tier 25,000 RPM · 2,000,000 TPM · 20,000,000 batch queue
Tier 35,000 RPM · 4,000,000 TPM · 40,000,000 batch queue
Tier 410,000 RPM · 10,000,000 TPM · 1,000,000,000 batch queue
Tier 530,000 RPM · 180,000,000 TPM · 15,000,000,000 batch queue

GPT-5.4 Mini is the mid-tier member of the GPT-5.4 generation — priced well under the full GPT-5.4 model, with a genuinely smaller context window rather than the same window at a discount.

That smaller window sits comfortably under the long-context repricing threshold that governs several larger models on this roster, and no repricing rule is published for this model at all — checked directly against its own page rather than assumed absent because the window looks small enough that it "probably doesn't apply." For any workload that fits inside this model's window in the first place, that's one less pricing mechanic to plan around compared to its larger siblings.

It shares its rate-limit group with the cheapest member of the current GPT-5.6 generation rather than with the rest of the GPT-5.4 family — worth checking directly rather than assuming rate limits track model generation the same way pricing does, since here they clearly don't.

Fast mode is published for this model, at a real premium over its already-low Standard rate, which makes it one of the more affordable ways to get lower latency on this site's entire roster even after the premium.

The table below is computed directly from this site's facts module: pricing across every published service tier and the rate-limit table.

Verified 2026-08-09 against https://developers.openai.com/api/docs/models/gpt-5.4-mini.

Could not confirm: No >272K long-context repricing rule is stated on this model's own page.

Checked: https://developers.openai.com/api/docs/models/gpt-5.4-mini · https://developers.openai.com/api/docs/pricing