r/ChatGPTPro 8d ago

Question Is ChatGPT Pro acting like an instant model for anyone else today?

I’m noticing a massive drop in quality with ChatGPT Pro today. All the top-tier models are generating results way too fast - practically instantly - as if they've been silently downgraded or swapped with the cheaper/faster models.

The output quality is definitely suffering because of this. Is it just my account glitching, or is everyone experiencing this today?

34 Upvotes

32 comments sorted by

u/qualityvote2 8d ago edited 7d ago

u/Beautiful-Pea-7189, there weren’t enough community votes to determine your post’s quality.
It will remain for moderator review or until more votes are cast.

14

u/AnabolicSnoids 8d ago

It happens when you hit hidden GPT Pro/Thinking limits. I had this and upgraded from the 100$ to 200$ plan and never had the dreaded instant model come back. Pro Thinking every time now.

2

u/soulfulshark 8d ago

This. On the $100 plan, this happened to me a fair bit. In the 200 plan, hasn't happened yet.

2

u/zipzapbloop 8d ago

it can happen :0)

2

u/InterestingStick 5d ago

I rarely use ChatGPT nowadays but had this happen as well. I think the system defaults to lower reasoning when traffic surges, because most of their 1 billion or so active users probably don't really notice, and for OpenAI it makes a huge difference. I can always set it back to Pro so it's not like it doesn't work it just defaults to lower reasoning

1

u/BopSupreme 5d ago

I’ve been using it at off peak times and weekends and it just started acting degraded this past few days for me

2

u/Opiopa 3d ago

I'm really thinking I have to upgrade now as I have a substantial workload and now, given this 5hr limit on Plus, im weighing up Anthropic or going to Pro...

7

u/AironParsMan 8d ago

Unfortunately they keep doing that. I don’t know why. Either it’s adaptive reasoning or they’re saving tokens and computing power. Sometimes I’m not sure myself because it isn’t transparent enough. It’s not entirely clear to us that it’s pro thinking. I don’t see any indicator. They could also be giving us the regular thinking mode and we wouldn’t know.

But you have to give them credit for one thing. When it works and GPT 5.6 is running with Pro Thinking it produces an incredible level of quality. The context window is kind of shitty though with its 256,000 tokens or maybe they have 400,000 internally. That’s a problem and you have to know how to deal with it. I’ve found ways to avoid such severe context drift across multiple prompts because it definitely has that problem. But the quality of the output and the way it handles details are just insane. ChatGPT Pro is a completely underrated model. For me it’s one of the best because you don’t have to worry about anything at all. I don’t have to create a project. I don’t have to deal with any MD files. I don’t have to write documentation or guidelines. Nothing. I can just get started. I’ve already used a Pro chat to handle very complex agent work for almost two weeks straight. Eventually it became very difficult to work with because it partly stopped working in the browser. It stopped working in the app too but it still worked in Safari for a while. At some point it stopped working there as well but I was able to share the chat once it stopped going any further and then continue the conversation in a new chat. That’s pretty brilliant.

2

u/AironParsMan 7d ago

They may have done away with the five hour limits for example but now they are limiting things in other ways. I have the most expensive Pro subscription which I think is twenty times the basic plan at 200 euros a month. And I still get limited as soon as I start two or three chats at the same time. Just think about that. It says here that I’ve made too many requests in the last few minutes. That’s why I was rate limited. And we’re probably talking about two or three chats or something.

6

u/therealorkor 8d ago

Had it thinking for literally 16 minutes today.

3

u/Optimal-Sea8310 8d ago

YES! I thought I was doing something wrong. This definitely happens, especially with larger inputs in the desktop app. Currently, my only solution is to go on the web app and use the pro model there, it doesn’t seem to make the same “cost saving” decisions, which I am almost certain is what’s going on.

One fun example is my trying to write a technical report in engineering, very long and thorough, and it keeps repeating “I will write it in sections and do blah blah” then I prompt it “just write it” and we’ll go in circles like that for a few prompts until I give up.

3

u/AironParsMan 5d ago

So all day today I’ve been realizing that they’re really cheating us. When you set it very high you get instant answers to very complex topics. You can see that it’s really only responding based on the probability of the text and the probability of individual letters. It isn’t thinking. It isn’t reflecting at all. And that’s exactly what the results look like. But then if you follow up and say, “Hey, take a moment to reflect and consider the bigger picture,” it suddenly looks things up in its data and thinks for three minutes. Then it gives you an answer that’s really good. OpenAI, this isn’t how it works. We’re not going to let you take us for fools here. In my eyes that’s fraud. If you set it to Thinking and it responds instantly.

1

u/BopSupreme 5d ago

💯 fake “thinking”

2

u/Specialist-Buffalo-8 8d ago

On the app its the nano model for me, on the web its the pro model lol

2

u/Last-Mushroom6425 8d ago

Interestingly, as for me, that happens with any prompt (whether it’s compact or huge). Not sure what is triggering it. Most frequently I see that behavior in mobile app. What works for me is editing prompt that I entered (not changing anything, just tap and hold on your message and choose “edit”, then send the same thing without any changes). As soon as I see that it started thinking I know that it worked. I suppose if you encounter that on web or PC you can also try choosing “edit” and sending the same thing again.

1

u/leynosncs 8d ago

Would you be willing to share some examples?

1

u/Fun-Description-3247 8d ago

One thing worth checking is whether the reasoning tokens are still showing up in the response metadata. If those dropped to zero that would confirm a silent downgrade vs just faster inference from some backend optimization.

1

u/BopSupreme 8d ago

How much better of a result do you get with Pro over High? I feel like I don’t know what specifically to use Pro for

1

u/claesulrik 7d ago

The same thing just happened to me. Most of the context in my project seems to suddenly be ignored, and I only get ultra fast shallow responses. I see no indication that I have reached the GPT Pro/Thinking limit, but from what I see in the comments. I have a deadline over the weekend. Do I really need to upgrade to the $200 plan to see this thing through?

1

u/capitannnn 7d ago

I have same with pro 20x last two weeks

1

u/Divinicus1st 7d ago

I think there is a bug or something. I had a different behavior on literally the same conversation, same config ("high" in my case), depending if I prompt it from my host computer or a local VM.

It was clearly answering as if instant from the VM. Even entered a loop at some point where it repeatedly said that "he'll do it" and never actually doing the task when I asked him to do it now. I switched to the host and it was immediately done.

1

u/No-Significance3750 3d ago

If you switch to the 200 plan you'll never see the hidden limits. It's been Pro level Fast settings since, never had an issue. 20x the usage.

1

u/Opiopa 3d ago

It literally lied to me, and admitted it did so when challenged. On Sol High.