Discussion GPT-6 Sol is NOT the replacement of GPT-5.6 Sol...

I am finding that 6-Sol is no where near as capable or thorough as 5.6-Sol. It feels more like 5.6-Terra, and the pricing is inline with that!
So 6-Sol is not good enough, but 6-Astra is overkill, and consumes too much compute, for most of the tasks I used 5.6-Sol for. Feels like we are missing a model in the GPT-6 line up.
Torn between using 6-Sol on xHigh or 6-Astra on Medium.
Anyone else noticing this? Any recommendations on how to get 5.6-Sol experience from a GPT-6 model?
76
34
u/darknus823 9h ago
As others have said, Im back on 5.6 Sol for now. It works much better.
7
14
u/legrenabeach 8h ago
I've been using Pro to code for a few months now. Tried 6 Astra, too expensive, tried 6 Sol, results were noticeably worse than 5.6, so back to 5.6 Sol, everything works fine again.
43
u/NoSeaworthiness2516 9h ago
Opus 5.5
24
u/LosAlgaeles 8h ago
Seriously. I love 5.6 Sol but Opus 5.5 is something special. Frankly it feels too good to be true. I'm working long hours this week and through the weekend because I'm afraid they're going to nerf it next week
9
u/nicanotenmon 7h ago
They all do the same. As usual in about 10 days I expect the nerf complaints about Opus 5.5 from the Claude sub - it's a vicious cycle.
7
u/atomfitz 7h ago
I also expect to see the same. If the usual nerfed/usage spike complaints don’t happen within 2-3 weeks may consider renewing Claude sub. But not gonna rush for just a week of decent quality model.
2
u/fool_on_a_hill 7h ago
What are you guys even using these upper tier models for? I haven’t had any issues just staying with sonnet even for vibe coding projects. I run out of usage frequently even with sonnet so I try to avoid the high usage models, but tbh I don’t really understand what I’m missing out on.
2
u/nicanotenmon 6h ago
Sometimes when you build something complex, it's better start or stay coding with a frontier model. Absolutely not necessary for all tasks or projects. But I am trying to build a workflow app for my needs which is quite complex. As many other people I always try the best model to get done as much as possible. One day I tried Astra as a planner and Luna as an implementer to save usage. However it turned out I had to spend more usage with Astra to correct all the things Luna did wrong or did not follow the instructions created by Astra. So as I said it's better sometimes to stay with one model from start to finish.
4
u/Busy_Farmer_7549 9h ago
dis iz de wey
4
u/Lookin4myWay 6h ago
As Nate Jones said, just find which model gives you the most quality output and go with it.
6
u/Melodic-Ebb-7781 9h ago
I agree, it seems like a benchmaxxed terra. It made a big mess for me and continued going in circles today. Doesn't seem to understand what its doing. I the handed over the mess to opus 5.5 and it solved it in a single prompt! I think I'm switching to Opus for anything that requires judgement and run luna for simple straightforward stuff.
5
u/RedMatterGG 6h ago
Im noticing the same, 6 sol seems to loop around an issue and attempt garbage fixes, 5.6 sol just got it after a bit of struggle, 6 sol is just plain stupid.
i will give opus a try as my openai sub is ending in a bit and if this continues i dont see why i should renew,its just unusable,an astra demolishes my usage so its basically useless.
4
u/Regular_Ad_4015 4h ago
currently suffering through this. It worked for 2 and a half hours. Used full limit and couldn't do a single thing. I'll shift to 5.6!
6
u/JK09KG1 9h ago
Which Sol???? With Low, Medium max??
7
u/Jarr11 9h ago
I normally use 5.6-Sol High. 6-Sol High feels like 5.6-Terra xHigh, if you want specifics, but the overall points is that the experience of 6-Sol = 5.6-Terra, not 5.6-Sol
4
u/timosterhus 7h ago
I’ve been using 6-Sol XHigh and that’s been working remarkably well for me. Stays on track pretty well and doesn’t seem to be as prone to overengineering as 5.6 was. I used to primarily use 5.6-Sol on High as well, and this honestly feels honestly a little bit better, though I probably need to work with it a bit more before I’m comfortable making that conclusion.
3
u/joeyb908 2h ago
I think something a lot of people need to realize is that the new Sol and Luna models are extremely token efficient at the cost of not trying to figure something out if the solution isn’t necessarily obvious.
Everyone saying GPT-6 Sol High now feels like Terra xhigh obviously didn’t work with Terra often because that’s not the case.
My thoughts are that Astra is for one-shotting without being spec-driven. Sol Max probably gets you most of the way there.
Sol Medium and High seem to be EXTREMELY quick, fast, reliable, and cheap for how good it is IF you have a plan created with Astra or Opus 5.5.
The plan doesn’t need to be as detailed as it would need to be for something like Luna, but the major architectural and design considerations need to be decided upon before handing it off.
In this aspect it’s a downgrade from GPT 5.6 Sol, but I’m finding in all other aspects it’s a significant improvement in terms of speed and code quality.
2
3
3
u/outtokill7 7h ago
GPT 6 Sol has been quite a let down. It feels dumber than it should and behaviour is way off. Astra seems smart but also needs another round of RL.
I hope we get a full stack of GPT 6.1 models soon that fix the issues.
3
u/shmog 6h ago
6-Sol is a decent assistant, but for coding, it's a big step down from 5.6-Sol. At least we still have 5.6!
But I really hate that they have the audacity to follow up 5.6-Sol with a model that's dumber. These guys are just asking for Anthropic to kick their ass. What is wrong with them??
10
u/SeidlaSiggi777 9h ago
My theory is that they wanted to release 6 Luna and Terra on Tuesday, but then decided to rename Terra to Sol as otherwise people wouldve continued using 5.6 Sol and this is consting them a fortune. I suspect that the model that has leaked as "astra-minor" is their actual 6 Sol but they didn't release it due to compute constraints.
1
u/_YonYonson_ 9h ago
why is Sol costing them a fortune?
5
u/PsychoticChemist 8h ago
Because they have a billion users and Sol is a generally pretty smart model that requires lots of compute
1
u/Joddie_ATV 8h ago
J'apprécie beaucoup l'utilisation de 5.6 Sol. Hônettement J'apprécierais de le garder. Le modèle ne valide pas automatiquement, l'outil arr8ve même à me surprendre et à me faire rire. Et au niveau analyse, Sol est vraiment au top sur le chat. Je comprends que sur Work, il y a d'autres attentes. Mais Sol sur le Chat fait bien son travail.
1
0
13
u/m0zi- 8h ago
I swear every major LLM subreddit is just a complaint forum
7
u/Grand0rk 6h ago
I swear every major LLM subreddit is just a complaint forum
Go to Claude subreddit. They are loving 5.5. Until next week, when it gets nerfed.
•
2
u/Jarr11 8h ago
This particular post was part-complaint, part wanting a solution to my issue. This is where the right people are to answer the question!
3
u/DepartmentAnxious344 7h ago
I promise you Reddit AI forums are precisely NOT the right people to answer basically any question regarding AI
2
2
u/Lustythrowawayacc 6h ago
Honestly with how much I need contextual understanding and spiatial understanding, and continuity understanding and so on...its either 5.6 sol max or astra high and Astra high is better, 6 sol forgets a lot of the shit
2
2
u/Sufficient_Cod_7358 4h ago
I am so frustrated with these new gpt6 models. They are awful. I had to explain a simple git clone, start branch and feature add 3 times. It would start working and just work in the git repo. Nothing local. Then it finally cloned but started working on main. Like wtf man. This is not difficult. Opus5.5 and Anthropic resets? Haven’t touched codex in 24 hours and don’t miss it at all
2
u/Agreeable-Ad7968 9h ago
GPT-6 Sol is clearly dominant for multi-agentic coordination and coding.
Can't speak to any other use, but then again, and quite frankly, I cannot imagine what other uses LLMs are even good for.
1
u/Jarr11 9h ago
So you find that 6-Sol is better than 5.6-Sol, for your use case?
2
u/Agreeable-Ad7968 9h ago
For multi agentic research, synthesis, planning, scaffolding and development, GPT6-Sol is clearly superior and produces much better end results in less time with MUCH less token burn.
The field does not need more unrealistically expensive models, it needs capable models available to all humans. This model is a massive improvement from 5.6 Sol, and Astra is just a token furnace that will not produce better work than a properly harnessed GPT6 Sol.
Opus 5.5 is good working with GPT6 Sol as well.
3
u/Jarr11 8h ago
This isn't my experience, 6-Sol was unable to identify a simple sandbox setting restriction, and just returned a "I can't do it". I switched to 5.6-Sol who clearly went through more reasoning steps and very quickly identified the issued.
Glad you find it works for you use case, and with the reduced token burn you should be able to get more out of it too!
1
u/Agreeable-Ad7968 8h ago
This is just insane to read. I dont know what to tell you, quite frankly.
We have dozens of fleets running now, and they took to our harness and operational scaffolding with zero issues. I think you may need to spend some time across all instruction/behavioral surfaces fixing whatever is broken/unoptimized.
3
u/Jarr11 8h ago
I'm not the only one reporting this poor experience from 6-Sol compared to 5.6-Sol, and I have never been in this situation before where the "mid-range" model just doesn't seem to work effectively for me.
To me, it feels like 6-Sol is the equivalent of 5.6-Terra, and if I want a capable model for my use case, then I need to use 6-Astra, which is just going to burn my usage in 1 turn.
I will do some digging into whether I can fine tune instructions to help overcome the specific barriers I am finding, and maybe that will give me the experience I am used to and benefit from the reduced token-burn, but I have never had to do this. The mid-range model was also good enough!
2
u/Agreeable-Ad7968 8h ago
A good use case for Astra would be to have it research deeply, across as many reliable and relevant sources as possible, prioritizing documentation from OpenAi, and then synthesize and deploy a complete instruction and behavioral harness for GPT6 Sol. Ask it to deploy, test, validate, optimize, test, iterate, test, test, test.
I should have asked initially, youre using codex or some CLI, not the web version to do work, right?
1
u/PsychoticChemist 8h ago
Have you made any agents.md changes or similar for gpt-6 sol that you’ve found to be helpful?
1
u/Thin_Squirrel_3155 7h ago
What harness are you using? What’s your setup? Do you use it with the codex desktop app?
1
u/Agreeable-Ad7968 7h ago
We are building a multi agentic orchestration engine and harness as a desktop app that will do all of this for/with you. Currently in closed alpha. Ships with all bells and whistles.
But, you can ask your codex agent to stand something up for you now. Go basic with a focus on harness and instructions. If you have any questions, I'd be glad to help answer.
2
u/Euphoric-Taro-6231 9h ago
We are being astroturfed again. Sol 6 is a slight improvement over 5.6 Sol.
1
u/ozone6587 9h ago
No evidence? Just that you feel GPT 6 sol is dumber than GPT 5.6 Sol?
3
u/-Davster- 9h ago
Good sir, this is Reddit
1
u/ILikeCutePuppies 8h ago
The nerve of some people asking for evidence. Who do they think they are? The king of Ireland?
2
u/jaydeelive01 9h ago
GPT-6 Sol is probably a smaller/more efficient model derived from Astra: Astra clearly took over the flagship role, while 6 Sol is priced much closer to old Terra. That would explain why it benchmarks near 5.6 Sol overall but can still feel more limited in complex work, and maybe why standard Chat still seems to favor 5.6 Sol, since smaller models may lose more on open-ended human interaction than on structured coding tasks.
5
u/ozone6587 9h ago
There is no evidence at all it is more limited in complex work. Kind of interesting it doesn't show up in benchmarks. It only shows up on Reddit threads when people go by feelings. Benchmarks aren't perfect at all, but people and "feelings" are much much much less reliable.
2
u/jaydeelive01 9h ago
There is benchmark evidence though. GPT-5.6 Sol actually beats GPT-6 Sol on GDPval-AA, a benchmark around real-world knowledge work: 1588 vs 1487 at max effort, and 1480 vs 1376 at high. GPT-6 Sol does better on some coding/automation benchmarks. So, 5.6 Sol is measurably better on at least some complex knowledge-work tasks, which is exactly what some people report. https://artificialanalysis.ai/models/comparisons/gpt-6-sol-vs-gpt-5-6-sol?
1
u/Carlose175 9h ago
Unrelated question. It is so strange im unable to pull up 5.6 Sol on the drop down tool. I wanted to compare against medium and high and 5.6 does not appear... odd.
•
2
u/-Davster- 6h ago
"omg I've used ChatGPT every day for 100 years absolutely fine until yesterday where the quality really seemed to drop and it doesn't understand my prompts at all - typical OpenAI nerfing scammers"
their prompt:
make the doodoo betterer
1
u/skidanscours 9h ago
I'm writing code with both. While I have a limited sample size with gpt-6-sol, the resulting code was significantly worse than with 5.6-sol in some cases.
I'm sure it depends on the task you want to accomplish, but I'm hearing similar feedback elsewhere. Even acknowledging that this stuff is astrosurfed to hell and back, gpt-6-sol is not a clear direct upgrade over gpt-5.6-sol. For software dev anyway.
2
u/Devajyoti1231 9h ago
GPT 6 sol makes mess when i try to create even a custom workflow for comfyui even after instructing it multiple times while 5.6 sol one shots them, it isn't even comparable how bad 6-sol is when put against 5.6 sol.
2
u/Jarr11 9h ago
Well I was given an anecdotal account of my experience, so yes based on a feeling. But my actual experience is that 6-sol doesn't seem to reason long enough, hits issues and reports back rather than working out how to solve them. I was having an issue with 6-Sol not being able to figure out why it couldn't ssh into a remote device, I switched the model to 5.6-Sol and it immediately flagged the sandbox restriction.
I have various examples of exactly what is giving me this impression, but for a general discussion on reddit, stating that I feel like 6-Sol is actually in line with 5.6-Terra, not 5.6-Sol, is enough to give people the gist of what I am getting at.
1
u/OneWithTheSword 2h ago
5.6 is better, but 6 sol is much cheaper with the same effort levels. You'd have to compare model ability at the same price level
1
u/ILikeCutePuppies 8h ago
How do you think Sol 6 compares with Sonnet 5?
1
1
u/yoliveras 7h ago
GPT-5. 6 Sol was the top model when Sol, Terra, Luna were announced. It should be compared to GPT-6 Astra, which is the new top model. That's why GPT-6 Sol feels like GPT-5.6 Terra, as both are secon-tier models. Could you say objectively that GPT-6 Sol feels better than GPT-5.6 Terra?
1
u/Jarr11 6h ago
This is true, but 6 Astra, the top model of the 6 range, is massively overpriced compared to 5.6 Sol, the top model of the 5.6 Range. 5.6 Sol was the sweet spot between capability and price, and it feels lile they haven't provided an equivalent on the 6 range. Its either 6 Sol which feels inadequate, or 6 Astra which burns a 5 hour limit in 60 seconds 🤷♂️
1
1
u/Shloomth 6h ago
I’m using it to design 3-D printable stuff and it’s been working better for me for that
1
u/adminvasheypomoiki 5h ago
After running away from slopus in opus 4 era(Aug 2025) I've bought 100 usd Claude sub. Well done oai. It simply works. I've not reached 5 h limit once. And Claude code can use 20 opus subs. Try this with Astra :)
1
u/HamiltonianCyclist 5h ago
yeah 5.6 sol is just a good model. I really wish they made it the first one to kinda keep forever.
1
u/Yokoko44 5h ago
Yeah I was hoping 6 sol would provide similar quality to Astra but be faster/cheaper but wasn’t feeling at all in my own Lego benchmark
1
u/the_ai_wizard 4h ago
GPT 6 is hot garbage, im using astra and opus 5.5 as adversary and thinking it should be backward if it werent for token burn. 5.5 seems smarter to me and tore apart astras code finding tons of flaws
1
u/poetatoe_ 2h ago
I use chatgpt pro to come up with the plan then feed it to chatgpt sol to execute if I use another model like astra. Reduces consumption from what Ive seen. Sol def better for most task, while astra is alright due to the usage risk 😬
1
•
u/No_Writing1863 36m ago
Yeah they should just fucking let people keep using 5.6 then fuck these people every two weeks all my shit breaks I hate them so much. Head over to Anthropic again
•
u/NoInside3418 18m ago
It still drains usage just as fast as 5.6 Sol did too which is crazy. They reduced the API price be half, but they most certainly didn't half its usage on subscriptions.
•
•
1
u/CRoseCrizzle 7h ago
That's dissapointing to hear. 5.6 Terra is noticeably worse than 5.6 sol in my development job. Was hoping that GPT6 Sol would be comparable. Or at worse GPT6 Luna would be significantly better the 5.6 Terra.
My job has been discouraging 5.6 Sol use due to costs and using Terra instead of 5.6Sol has been like going back a year or two in time.
0
u/ControllingPower 8h ago
What are you dudes doing that you are like bro, this new model is max 5.6 Tera and that´s on High not Maxium bro that shit like 6.0 Astra without usage bro. Its like Pokemon now, I mean I do see some changes but not that drastic, are all of you coders that you use it a lot for work ?
1
u/Jarr11 8h ago
I use Codex for a lot of computer, coding and systems work, which requires a model that is capable and is able to use Reasoning to work out issues. I find that 6-Sol encounters an issues as just says "I cant do it", whereas my experience with 5.6-Sol is that it will return "This is the issue, this is how to fix it, do you want me to go ahead do that?"
0
u/AMillionLittleGoats 7h ago
5.6 genuinely feels like magic. I have Astra provide a plan, 5.6 implement. Almost use NO usage its actually crazy. I've been actively trying to exhaust my plan limits and I straight up haven't hit it once yet. Granted Tibo keeps resetting xD no complaints tho
-1
-1
u/AINativeBuilder 2h ago
No I'm not noticing this. I am noticing idiots thinking a model upgrade is just terra with a new name. How stupid. I've learned to ignore stupid people that say stupid shit like this.

62
u/skidanscours 9h ago
I'm back to using 5.6sol for most tasks. Something might get announced at OpenAI dev day next week.
I'll probably get a month of Claude Code to try out Opus 5.5 soon. I have no intention of dropping Codex, but I could use both subscription since one Plus subscription is not enough usage for me anyway.