Verdict Is It Worth It

GPT-5.6 Sol's Effort Slider: Worth Touching, or Leave It on Auto?

OpenAI gave Plus and Pro users a dial for how hard ChatGPT thinks. I ran it through a week of real work to find out when it's worth moving.

Hands typing on a laptop keyboard with a second laptop alongside
Photo via Unsplash
The receipts
  • August 6: OpenAI shipped an updated GPT-5.6 Sol and Luna in ChatGPT with a four-position effort slider — Instant, Medium, High, Extra High — for Plus and Pro.
  • OpenAI's own safety card claims a ~60% cut in factual error rate versus GPT-5.5 Instant across its three test prompt sets, plus a +15.6 point jump on HealthBench Professional.
  • Extra High is gated to Pro, Business, and Enterprise — Plus tops out at High, so the headline mode isn't the one most people are paying for.
  • Reddit's read on Sol since its July debut has been split: cheaper and faster at code, still losing UI/design matchups to Claude Fable 5.
Short answer

GPT-5.6 Sol's effort slider (Instant, Medium, High, Extra High), shipped August 6, 2026, lets ChatGPT Plus and Pro users pick how much reasoning effort a response gets. OpenAI reports roughly a 60% drop in factual errors versus GPT-5.5 Instant and a 15.6-point HealthBench Professional gain. Extra High requires Pro or higher; Plus caps at High.

A slider. That’s the ship.

OpenAI’s August 6th update to ChatGPT didn’t rename anything or drop a new model number on you. It gave Plus and Pro users a dial — Instant, Medium, High, Extra High — for how much thinking GPT-5.6 Sol does before it answers. I’ve been running it against real work for a week: client research, a Python bug that wasn’t obvious, a handful of throwaway questions I’d normally just fire at Instant. Here’s where I land.

The receipts are real, and they’re OpenAI’s own

Per OpenAI’s deployment safety card for the August build, GPT-5.6 Sol cut its factual error rate roughly 60% across all three of its internal prompt sets compared to GPT-5.5 Instant, and picked up +15.6 points on HealthBench Professional. Both the Sol and Luna August builds are now rated “High” capability in the biological/chemical and cybersecurity domains under OpenAI’s own framework — worth knowing if you’re using this for anything adjacent to those.

I’m not independently re-running HealthBench. But the day-to-day pattern matched the claim: bumping the slider to High on a multi-step research task caught a source I’d have missed on Instant, and it did it without me writing a longer prompt. That’s the actual value of a slider over a fixed model — you’re not upgrading your writing, you’re upgrading the model’s patience on the same input.

Extra High is the mode you’re not getting

Here’s the catch nobody’s leading with: Extra High — the top of the dial, the one that pairs with the benchmark numbers — needs Pro, Business, or Enterprise. Plus subscribers get the slider UI but cap at High. If you’re paying $20/mo and reaching for this expecting the full jump, you’re testing a narrower band than the headline suggests. Worth reading our ChatGPT Plus breakdown before assuming the upgrade lands in your tier.

Where Sol still loses

None of this erases what Reddit’s been saying about Sol since it launched in July. The consistent thread across r/OpenAI and r/codex: Sol is the cheaper, faster coder — one widely cited benchmark ran $8.39 on Sol against $21.63 on Fable 5 for comparable output. But the same crowd keeps handing design and UI work back to Claude — see how Fable 5 stacks up against Sol directly if that’s your workload. A slider for reasoning effort doesn’t touch that gap. It’s not trying to.

The verdict

If your work involves anything you’d manually fact-check anyway — planning, research, debugging — move the slider up and let it earn the extra seconds. If you’re doing quick lookups or casual chat, Instant and Medium are still the right default; there’s no reason to burn effort budget on “what’s the capital of Peru.” And if you’re on Plus expecting Extra High, you’re not getting it — that’s a Pro line item, not a Plus feature.

The accuracy gains are legitimate and OpenAI showed its work. The gating is the part I’d actually complain about.

#TheAIMogul

Bottom lineMove the slider for anything you'd actually double-check by hand — planning, research, a gnarly bug. Leave it on Instant or Medium for everyday chat. The accuracy gains are real and OpenAI's own numbers back them, but Extra High is a Pro-tier toy, not a Plus feature, so most of you are testing High and calling it a day.

Frequently asked

What is the GPT-5.6 Sol effort slider?
A control added to ChatGPT on web, mobile, and desktop on August 6, 2026, letting Plus and Pro users choose Instant, Medium, High, or Extra High reasoning effort per response, instead of ChatGPT picking a fixed depth for every message.
Do I need ChatGPT Pro to use it?
You need Pro, Business, or Enterprise to reach Extra High. Plus subscribers get the slider too, but it caps at High — so the deepest reasoning mode isn't included in the $20/mo tier.
Is GPT-5.6 Sol actually more accurate?
OpenAI's deployment safety documentation reports about a 60% reduction in factual error rate versus GPT-5.5 Instant across its three internal prompt sets, and a 15.6-point improvement on HealthBench Professional. I haven't independently re-run those benchmarks; that's OpenAI's own claim.
How does it compare to Claude Fable 5 for coding?
Reddit threads tracked since Sol's July launch peg it as cheaper and faster on straight coding tasks, with one widely cited benchmark run costing $8.39 on Sol against $21.63 on Fable 5. The same threads say Fable 5 still wins on UI and visual design output.
What happened to free and Go users?
OpenAI moved free and Go tiers to a new default model as part of the same August 6 update, separate from the Sol effort slider rollout for paid tiers.