GPT-5.4 Mini — Specs, Pricing & Limits
Specs
| Status | Priced |
|---|---|
| Context window | 400,000 tokens |
| Max output | 128,000 tokens |
| Knowledge cutoff | Aug 31, 2025 |
| Reasoning effort levels | Not published |
| Long-context rule | Not published |
Pricing (short context)
| Standard — input | $0.750/MTok |
|---|---|
| Standard — cached input | $0.075/MTok |
| Standard — cache writes | Not published |
| Standard — output | $4.50/MTok |
| Fast — input | $1.50/MTok |
| Fast — output | $9.00/MTok |
Rate limits by usage tier
| Free | Not supported |
|---|---|
| Tier 1 | 500 RPM · 500,000 TPM · 5,000,000 batch queue |
| Tier 2 | 5,000 RPM · 2,000,000 TPM · 20,000,000 batch queue |
| Tier 3 | 5,000 RPM · 4,000,000 TPM · 40,000,000 batch queue |
| Tier 4 | 10,000 RPM · 10,000,000 TPM · 1,000,000,000 batch queue |
| Tier 5 | 30,000 RPM · 180,000,000 TPM · 15,000,000,000 batch queue |
GPT-5.4 Mini is the mid-tier member of the GPT-5.4 generation — priced well under the full GPT-5.4 model, with a genuinely smaller context window rather than the same window at a discount.
That smaller window sits comfortably under the long-context repricing threshold that governs several larger models on this roster, and no repricing rule is published for this model at all — checked directly against its own page rather than assumed absent because the window looks small enough that it "probably doesn't apply." For any workload that fits inside this model's window in the first place, that's one less pricing mechanic to plan around compared to its larger siblings.
It shares its rate-limit group with the cheapest member of the current GPT-5.6 generation rather than with the rest of the GPT-5.4 family — worth checking directly rather than assuming rate limits track model generation the same way pricing does, since here they clearly don't.
Fast mode is published for this model, at a real premium over its already-low Standard rate, which makes it one of the more affordable ways to get lower latency on this site's entire roster even after the premium.
The table below is computed directly from this site's facts module: pricing across every published service tier and the rate-limit table.
Verified 2026-08-09 against https://developers.openai.com/api/docs/models/gpt-5.4-mini.
Could not confirm: No >272K long-context repricing rule is stated on this model's own page.
Checked: https://developers.openai.com/api/docs/models/gpt-5.4-mini · https://developers.openai.com/api/docs/pricing