GPT-5.6 Terra — Specs, Pricing & Limits
Specs
| Status | Priced |
|---|---|
| Context window | 1,050,000 tokens |
| Max output | 128,000 tokens |
| Knowledge cutoff | Feb 16, 2026 |
| Reasoning effort levels | Not published |
| Long-context rule | >272,000 input tokens: 2x input, 1.5x output (full request) |
Pricing (short context)
| Standard — input | $2.00/MTok |
|---|---|
| Standard — cached input | $0.200/MTok |
| Standard — cache writes | $2.50/MTok |
| Standard — output | $12.00/MTok |
| Fast — input | $4.00/MTok |
| Fast — output | $24.00/MTok |
Rate limits by usage tier
| Free | Not supported |
|---|---|
| Tier 1 | 500 RPM · 500,000 TPM · 1,500,000 batch queue |
| Tier 2 | 5,000 RPM · 1,000,000 TPM · 3,000,000 batch queue |
| Tier 3 | 5,000 RPM · 2,000,000 TPM · 100,000,000 batch queue |
| Tier 4 | 10,000 RPM · 4,000,000 TPM · 200,000,000 batch queue |
| Tier 5 | 15,000 RPM · 40,000,000 TPM · 15,000,000,000 batch queue |
GPT-5.6 Terra sits in the middle of its family — priced well under Sol, well above Luna — and it's worth knowing exactly what that middle position does and doesn't buy you, because it isn't a straightforward "less capable, less expensive" tradeoff on every axis.
Context window, maximum output length and knowledge cutoff are all identical to Sol's — Terra doesn't trade any of those away for its lower price. What actually differs is the per-token rate itself, at a meaningful discount to Sol across every service tier and both the short- and long-context bands, and Terra also shares Sol's exact rate-limit group rather than sitting in a cheaper, more restricted one of its own. That combination — same window, same throughput tier, lower price — makes Terra worth checking seriously for any workload currently defaulting to Sol purely because Sol is what the CLI ships with, rather than because a session specifically needs Sol's particular strengths.
The published long-context repricing rule applies to Terra the same way it applies to Sol: crossing the threshold reprices the entire request, not just the excess, so a workload planned against Terra's lower base rate still needs to respect the same cliff Sol's own page describes.
The table below is computed directly from this site's facts module — pricing across every published service tier, the long-context rule, and the published rate-limit table.
Verified 2026-08-09 against https://developers.openai.com/api/docs/models/gpt-5.6-terra.
Checked: https://developers.openai.com/api/docs/models/gpt-5.6-terra · https://developers.openai.com/api/docs/pricing