Migrating: GPT-6 Sol to GPT-6.1 Sol
Green marks the cheaper/newer side per row where a direction genuinely applies
| Context window | 1,050,000 tokens | 1,050,000 tokens |
|---|---|---|
| Knowledge cutoff | Apr 20, 2026 | Apr 30, 2026 |
| Standard input price | $2.00/MTok | $2.00/MTok |
| Standard cached input | $0.20/MTok | $0.10/MTok |
| Standard output price | $10.00/MTok | $10.00/MTok |
| Cache writes | Surcharged (write rate) | Surcharged (write rate) |
| Reasoning effort levels | none, low, medium (default), high, xhigh, max | low, medium (default), high, xhigh, max |
| Long-context repricing rule | Published | Published |
| Rate-limit tier group (Tier 1) | 500 RPM | 500 RPM |
| Codex with ChatGPT sign-in | No retirement announced | No retirement announced |
OpenAI specifically flags this migration: if you already use gpt-6-sol, it says to review its migration guidance before switching. That's notable for a move between two models released a week apart at the same input and output price, and the reason is the effort floor.
The change that breaks requests
GPT-6 Sol accepts none reasoning effort; GPT-6.1 Sol doesn't, and doesn't accept minimal either. Requests that set either need low instead. Two knock-on effects follow. Reasoning tokens start appearing where there were none, billed as output, so a latency-tuned call gets a little slower and a little dearer per request. And on the API, the sampling parameters that only work with none — temperature, top_p and log probabilities — have to come out. Function calling changes too: GPT-6 Sol supports it in Chat Completions only with effort set to none, while GPT-6.1 Sol requires the Responses API for any tool calling.
The change that saves money
Cached input. GPT-6.1 Sol bills cache reads at 0.05x its input rate instead of 0.1x. Input, output and cache-write rates are unchanged, so every request is the same price or cheaper, and cache-heavy agentic sessions get noticeably cheaper.
What carries over
Context window, maximum output, long-context rule, and request and token limits per minute. GPT-6.1 Sol's knowledge cutoff is slightly newer. Its rate-limit table has no batch-queue column, so if you plan Batch capacity, note that no published queue limit exists for it.
Before switching
If your traffic never sets none, this is close to a free upgrade — and it's the CLI's default. If it does, test the same tasks at low and check whether the added latency and output tokens are acceptable before moving. GPT-6.1 Sol vs GPT-6 Sol has the full comparison.
Verified 2026-10-01 against CodexHow facts module (src/data/facts/) — see /about/#accuracy.