CodexHowSupport Us

GPT-6.1 Sol vs GPT-5.6 Sol

Green marks the cheaper side per row — informational only, never a verdict for your workload

GPT-6.1 SolGPT-5.6 Sol
Statuspriceablepriceable
Context window1,050,000 tokens1,050,000 tokens
Knowledge cutoffApr 30, 2026Feb 16, 2026
Standard input$2.00/MTok$4.00/MTok
Standard cached input$0.10/MTok$0.40/MTok
Standard output$10.00/MTok$20.00/MTok
Cache writesWrite rate $2.50/MTokWrite rate $5.00/MTok
Reasoning effort levelslow, medium (default), high, xhigh, maxnone, low, medium (default), high, xhigh, max
Faster tiers offeredFastFast
Long-context repricing rulePublishedPublished
Tier 1 rate limit500 RPM · 500,000 TPM500 RPM · 500,000 TPM
ChatGPT credits per 1M (in / out)50 / 250100 / 500
Codex with ChatGPT sign-inNo retirement announcedNo retirement announced

GPT-5.6 Sol: promotional pricing — OpenAI lowered GPT-5.6 Sol's prices on 2026-08-21 and says the promotional pricing is available at least through 2026-11-21. In ChatGPT plans it applies to usage paid with purchased credits; included plan usage and the 5-hour and weekly limits did not change.

This is the old Codex CLI default against the new one. GPT-6.1 Sol replaced GPT-5.6 Sol as the model the CLI runs when you don't choose, so any setup that never pinned a model has already made this switch — and any setup that pinned gpt-5.6-sol, or the unsuffixed gpt-5.6 alias, hasn't.

What the newer model wins

Price, on every cell. GPT-6.1 Sol is cheaper per input token, per cached input token, per cache write and per output token — and that comparison is against GPT-5.6 Sol's promotional rates, which OpenAI says are available at least through November 21, 2026. Its knowledge cutoff is a couple of months newer. And its cached-input discount is the deepest of any current model, which compounds in long agentic sessions.

What the older model keeps

The none effort level. GPT-5.6 Sol accepts everything from none to max; GPT-6.1 Sol starts at low. A workload that switches reasoning off for speed can't carry that setting across. GPT-5.6 Sol also publishes a batch-queue limit per usage tier, which GPT-6.1 Sol's table omits. And it has a track record on your workload, which is worth something if your prompts are tuned to it.

What stays the same

The context window, maximum output, long-context threshold and request and token limits per minute all match. Both offer Standard, Batch, Flex and Fast.

How to decide

Unless you depend on none effort, the newer default is cheaper on every row and newer on recency. Test a representative task on both before moving a tuned workload, and see the migration page for what to change in config.

Verified 2026-10-01 against CodexHow facts module (src/data/facts/) — see /about/#accuracy.