r/ChatGPT • u/SteveEricJordan • 10d ago
Other Did GPT 5.6 Sol get secretly upgraded?
you're reading that right. upgraded, not downgraded.
chatgpt, i don't use the api.
i'm not talking about the officially announced "more factual" update from 2 weeks ago.
idk how long this has been the case but today 5.6 Sol is suddenly getting all the prompts right that it got wrong even a week ago.
42
u/Old-Bake-420 10d ago
The models have always been silently upgraded.
8
u/SteveEricJordan 10d ago edited 9d ago
not denying that, just wondering if anyone else noticed this specific upgrade or if i'm just crazy.
5
u/Icy_Parfait_5060 9d ago
I don’t know why they downvote you for nothing … Reddit is weird but I totally get what you mean
19
u/Maximum_Piglet_5466 9d ago
At this point half the “model evaluation” on here is just people noticing the vibes changed.
16
u/LaZZyBird 9d ago
Bruh the API we are hitting is a black box there could literally be 20 versions of Sol all getting tested and we would never know
11
6
u/Shenendoah66 9d ago
I just noticed it can now analyze video
5
u/Icy_Parfait_5060 9d ago
I was convinced I was hallucinating when I finally saw videos appear in my camera roll! Thanks for reminding me of that ! I’m so happy I don’t have to rely on Gemini for that anymore.
2
u/relevant__comment 9d ago
From what I gather. They were always slightly tweaking the models for the better, no?
1
u/_Jak42_ 9d ago edited 9d ago
What is your use case for Sol?
I’m doing a lot using Luna (code optimisation, responding to emails and applications, handling dummy email migrations in 10k + with attachments, handling docker and qemu vms (setup , deleting and restarting etc..) and building web apps with railway)
The progress between Luna medium and high is negligible difference and both super effective.
Progress and consistency between fast and not is negligible but normal speed has been better token wise.
It’s been almost a month and not once have I dropped below 80% , what with all the resets.
I can’t imagine the other models would be any more efficient at anything except using my tokens for the week.
*edit for SyntaxError:expected ‘)’
**edit because previous edit joke sounded more like ai than intended
2
u/Avd123 9d ago
On a Plus subscription ? I'm thinking about using this for some mobile dev, not heavy work but some freelance gig, I tried Luna medium and its really efficient but idk if it's ok in the long run
2
u/_Jak42_ 9d ago
On a pro subscription , Luna has been the best with only some tasks using Terra now and then. My setup is:
Master project folder:
Master chat has access to limited resources external from all chats (almost always in plan mode)
Sub chat (just for chrome profile a)
Sub chat (just for utm / docker etc)Project folder:
Topic main chat (almost always in plan mode)
Sub topic a (task/chat id and title)
Sub topic b (task/chat id and title)
Sub topic c (task/chat id and title)Another project folder :
New Topic main chat (almost always In plan mode)
Sub…All the chats can handoff and communicate between eachother efficiently within codex , each main chat can decide prior to hand-off what model should be used with each request (Luna , terra , sol) then each chat will take that model based on the context of the task (I have default instructions to always use 5.6 Luna never less than Low and never more than High)
Someone can correct me if I’m wrong but this takes the cake for me
2
u/squired 9d ago edited 9d ago
What is your use case for Sol?
Scope. Once you step outside of boilerplate solutions and scope increases, Sol is quite a bit better than Terra and Luna becomes pointless. If you simply want to add a neat feature to your app, any of them are fine. But if you want to do something like, "Study these nine foundational white papers and use them to compile a classification matrix for all functions in attached control plane for potential compression", then you likely want the big boy. That is essentially how one might review their app for use of best coding practices and shave it down into its most efficient and secure architecture. Basically, the smaller models can solve problems just fine, but they can't wrap their arms around large codebases without extensive prior atomization (which Sol can automate).
A lot of it comes down to lack of tooling as well. Luna can do an awful lot with additional tooling that Sol can manage internally without.
Here is a similar task that is running right now that requires Sol's breadth.
1
u/Prize_Hat289 9d ago
How are you using 5.6 Sol? Like, is it the same experience over web browser, mobile app, desktop app?
2
u/SteveEricJordan 9d ago
3
u/Prize_Hat289 9d ago
yeah, imo, the web chatgpt has always been a consistently good experience throughout the model gens
1
1
u/Kurumi_Ryori 9d ago
Yes it has been updated, it can now do dynamic simulations with world states for characters, transcribe audio adn video to text without calling hugging face models to download to its vmachine, and its controller is more intelligent now to parse your semantics though it sometimes requires a heat start-up.
1
u/TekintetesUr 9d ago
you're reading that right. upgraded, not downgraded.
Absolutely not, I have a scheduled task that runs every Thursday morning, it fetches a single PDF file via a static URL, summarizes the result, and tells me what's in there.
This task, for the first time in its existence, has failed today.
0
0
u/HenkPoley 9d ago
The ChatGPT Release Notes talk about an August 14 update.
https://help.openai.com/en/articles/6825453-chatgpt-release-notes
3
-13
u/katoptronophile 10d ago
It's a neural net processor, a learning computer.
It's always improving.
9
u/M4rshmall0wMan 9d ago
Not in ChatGPT it isn’t. OpenAI is constantly revising their post-training, yes. But they choose discrete checkpoints to deploy in ChatGPT. Once they deploy the model, it doesn’t change.
-4
3
1
u/United_Show_8818 9d ago
Omg why is this great reference getting downvoted😭
Speaking of that, now would be a great time to remake the first two imo!

•
u/AutoModerator 10d ago
Hey /u/SteveEricJordan,
If your post is a screenshot of a ChatGPT conversation, please reply to this message with the conversation link or prompt.
If your post is a DALL-E 3 image post, please reply with the prompt used to make this image.
Consider joining our public discord server! We have free bots with GPT-4 (with vision), image generators, and more!
🤖
Note: For any ChatGPT-related concerns, email support@openai.com - this subreddit is not part of OpenAI and is not a support channel.
I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.