CodexHowSupport Us

GPT-5.6 Terra — Specs, Pricing & Limits

Specs

StatusPriced
Context window1,050,000 tokens
Max output128,000 tokens
Knowledge cutoffFeb 16, 2026
Reasoning effort levelsNot published
Long-context rule>272,000 input tokens: 2x input, 1.5x output (full request)

Pricing (short context)

Standard — input$2.00/MTok
Standard — cached input$0.200/MTok
Standard — cache writes$2.50/MTok
Standard — output$12.00/MTok
Fast — input$4.00/MTok
Fast — output$24.00/MTok

Rate limits by usage tier

FreeNot supported
Tier 1500 RPM · 500,000 TPM · 1,500,000 batch queue
Tier 25,000 RPM · 1,000,000 TPM · 3,000,000 batch queue
Tier 35,000 RPM · 2,000,000 TPM · 100,000,000 batch queue
Tier 410,000 RPM · 4,000,000 TPM · 200,000,000 batch queue
Tier 515,000 RPM · 40,000,000 TPM · 15,000,000,000 batch queue

GPT-5.6 Terra sits in the middle of its family — priced well under Sol, well above Luna — and it's worth knowing exactly what that middle position does and doesn't buy you, because it isn't a straightforward "less capable, less expensive" tradeoff on every axis.

Context window, maximum output length and knowledge cutoff are all identical to Sol's — Terra doesn't trade any of those away for its lower price. What actually differs is the per-token rate itself, at a meaningful discount to Sol across every service tier and both the short- and long-context bands, and Terra also shares Sol's exact rate-limit group rather than sitting in a cheaper, more restricted one of its own. That combination — same window, same throughput tier, lower price — makes Terra worth checking seriously for any workload currently defaulting to Sol purely because Sol is what the CLI ships with, rather than because a session specifically needs Sol's particular strengths.

The published long-context repricing rule applies to Terra the same way it applies to Sol: crossing the threshold reprices the entire request, not just the excess, so a workload planned against Terra's lower base rate still needs to respect the same cliff Sol's own page describes.

The table below is computed directly from this site's facts module — pricing across every published service tier, the long-context rule, and the published rate-limit table.

Verified 2026-08-09 against https://developers.openai.com/api/docs/models/gpt-5.6-terra.

Checked: https://developers.openai.com/api/docs/models/gpt-5.6-terra · https://developers.openai.com/api/docs/pricing