Potential silent model rerouting: GPT-5.6 Pro behaves as Instant / GPT-5.5 Mini
Summary
Selecting GPT-5.6 Pro does not engage the expected Pro/reasoning behavior. The response is returned immediately with no visible thinking, behaves like Instant, and when asked which model is responding, it identifies itself as GPT-5.5 Mini.
This is consistently reproducible across multiple surfaces on the same account:
- ChatGPT Web on Windows, including a private/incognito browser session
- Codex Desktop after switching to the ChatGPT/Work surface
- ChatGPT for iOS on an iPhone 16 Pro Max
The ChatGPT app and iOS are both updated to the latest available versions.
Steps to reproduce
- Start a new chat.
- Select GPT-5.6 Pro.
- Send a prompt that would normally engage Pro reasoning.
- Observe that the response appears immediately, with no reasoning/thinking phase.
- Ask which model is responding; it reports GPT-5.5 Mini.
- Repeat in an incognito browser session and on iOS.
Expected behavior
The request should be served by GPT-5.6 Pro when that model is explicitly selected.
If the service must reroute or downgrade the request, the UI should clearly disclose the actual served model and explain how the user can resolve the restriction.
Actual behavior
Every attempt after selecting Pro behaves as Instant: no thinking phase and an immediate response. The same result occurs across Windows Web, Codex Desktop, and iOS.
The model's self-identification is not treated as conclusive backend proof by itself. However, the consistent absence of Pro reasoning and apparent automatic Instant behavior across multiple clients suggests a possible requested-model versus served-model routing mismatch.
Request
Please inspect the server-side routing for the affected account/session, including the requested model, served model, routing policy, and any account-side eligibility or safety flag. Please confirm whether this downgrade is expected and restore normal GPT-5.6 Pro access if it is a false positive.
This may be related in pattern to #11189 and #11971, although the affected models and client surfaces are different.
10 Comments
Potential duplicates detected. Please review them and close your issue if it is a duplicate.
Powered by Codex Action
i also encounter this. when asked what model are you, the response is im 5.5-mini
Me too. Wonder when it can be fixed...or whether they're going to fix it
Thank God someone brought this up.
I feel like this has been happening for almost a month now — the Thinking mode is completely gone from my GPT chat window.
Yes, I’m on the free plan, so I know there are still limits. But I really miss how useful it used to be.
This is what my interface looks like. You can see that the Thinking mode option is simply not there anymore.
<img width="1067" height="518" alt="Image" src="https://github.com/user-attachments/assets/95beeb9a-1db5-402c-a1ab-6a1033cd1af5" />
I’m starting to believe that this issue may be related to the IP address being flagged. It most likely occurs among users in China who are using VPNs.
Is that so, bro? If that’s the case, then I guess there’s nothing we can do about it.
That said, there is a workaround: press F12, enable the mobile view in the developer tools, and you’ll be able to select Thinking mode again.
Once you do that, even after switching back to the desktop interface, the blue “Think” tag will stay there and can still be copied.
I reproduced a superficially similar symptom on ChatGPT Web in Safari (macOS, ChatGPT Pro, 2026-07-29), but a controlled check did not establish silent rerouting:
This is evidence of a reproducible model self-identification / unsupported routing-metadata claim issue, not proof of a backend model downgrade. It matches the distinction described in #33838: maintainers should verify requested/served model from server-side telemetry rather than treating the assistant’s self-report as authoritative.
No conversation ID, account identifier, raw network log, token, or other private diagnostic data is included here.
Follow-up after correcting the test methodology: the earlier
OK-style control was too simple and should not have been used to rule out a downgrade.I repeated the comparison in fresh ChatGPT Web conversations using the same nontrivial digit-DP prompt and a locally verified answer:
In this controlled sample, Pro was actually faster than Instant. Therefore response speed—even a very fast response—is not a reliable identifier of the served model. Simple identity prompts in the originally reported conversation completed in roughly 2–5 seconds, but those prompts do not require Pro-level reasoning and cannot distinguish the route.
The user's report of frequent Instant-like behavior after selecting Pro remains unresolved. The strongest actionable request is still for OpenAI to inspect server-side telemetry for requested model, served model, and any fallback/routing reason. The current picker state is mutable and does not establish which model served an earlier turn.
No private conversation identifiers or raw diagnostics are included in this public follow-up.
I have a related but distinct HAR-verified reproduction. Earlier I observed GPT-5.6 Pro falling back to GPT-5.5-mini; later, GPT-5.6 Thinking/Sol independently began falling back in Chrome.
The later phase has a strict same-account, same-workspace, same-conversation comparison:
No limit/capacity notice appeared. The sanitized HAR files were sent privately under Support Case
13542447and are not being posted publicly.Because Pro fallback and the later Thinking/Sol fallback appeared at different times, I do not think they should be assumed to have the same cause without comparing internal routing logs.
Is this issue going on for so long? Users from UK started expereinceing this in the last 4 days since 20.08.2026. USing GPT 5.6 Sol frontend but backend resolves to GPT 5.5 mini and even VPn cant help on it. Any chance to fix this as this is a disaster!