GPT-6 Sol — Specs, Pricing & Limits
Specs
| Status | Priced |
|---|---|
| Released (API changelog) | 2026-09-22 |
| Context window | 1,050,000 tokens |
| Max output | 128,000 tokens |
| Knowledge cutoff | Apr 20, 2026 |
| Reasoning effort levels | none, low, medium (default), high, xhigh, max |
| Long-context rule | >272,000 input tokens: 2x input and cache rates, 1.5x output (full request) |
| Cache writes | Billed at the cache-write rate instead of the input rate |
| ChatGPT credit rate (per 1M) | 50 in · 5 cached · 250 out |
Pricing per 1M tokens, every published service tier
| Standard — short context | in $2.00 · cached $0.20 · writes $2.50 · out $10.00 |
|---|---|
| Standard — long context | in $4.00 · cached $0.40 · writes $5.00 · out $15.00 |
| Batch — short context | in $1.00 · cached $0.10 · writes $1.25 · out $5.00 |
| Batch — long context | in $2.00 · cached $0.20 · writes $2.50 · out $7.50 |
| Flex — short context | in $1.00 · cached $0.10 · writes $1.25 · out $5.00 |
| Flex — long context | in $2.00 · cached $0.20 · writes $2.50 · out $7.50 |
| Fast — short context | in $4.00 · cached $0.40 · writes $5.00 · out $20.00 |
| Fast — long context | in $8.00 · cached $0.80 · writes $10.00 · out $30.00 |
Rate limits by usage tier
| Free | Not supported |
|---|---|
| Tier 1 | 500 RPM · 500,000 TPM · 1,500,000 batch queue |
| Tier 2 | 5,000 RPM · 1,000,000 TPM · 3,000,000 batch queue |
| Tier 3 | 5,000 RPM · 2,000,000 TPM · 100,000,000 batch queue |
| Tier 4 | 10,000 RPM · 4,000,000 TPM · 200,000,000 batch queue |
| Tier 5 | 15,000 RPM · 40,000,000 TPM · 15,000,000,000 batch queue |
GPT-6 Sol is the model OpenAI names as GPT-5.5's replacement in Codex on paid ChatGPT plans, which makes it the most likely destination for anyone working through the GPT-5.5 retirement. Days after it shipped, GPT-6.1 Sol arrived alongside it and took over as the CLI's default, so GPT-6 Sol now sits in an odd position: current, recommended for a specific migration, and already succeeded.
Against GPT-6.1 Sol the price table is close. Input and output cost the same; cached input costs more on GPT-6 Sol, where reads bill at the usual 0.1x of input. The context window and output ceiling match, and its knowledge cutoff is slightly earlier.
Its practical advantage is the effort floor. GPT-6 Sol still accepts none, from none up to max with medium as the default, so a latency-sensitive workload that depends on switching reasoning off can move here without changing that setting — something GPT-6.1 Sol and GPT-6 Astra won't allow. OpenAI also notes that GPT-6 Sol supports function calling in Chat Completions only with reasoning effort set to none; with reasoning on, use the Responses API for tools. Fast mode isn't available for it with EU data residency.
Its rate-limit table is the same Sol table GPT-5.5 used, batch-queue column included, so throughput doesn't change on the retirement migration. GPT-6 Sol vs GPT-5.5 shows that move in full, and GPT-6.1 Sol vs GPT-6 Sol the choice between the two current Sol models.
The tables above are computed from this site's facts module: every published service tier with both context bands, the effort levels, and the rate-limit table.
Verified 2026-10-01 against https://developers.openai.com/api/docs/models/gpt-6-sol.
Checked: https://developers.openai.com/api/docs/models/gpt-6-sol · https://developers.openai.com/api/docs/pricing · https://developers.openai.com/api/docs/changelog · https://learn.chatgpt.com/docs/models