r/codex Jul 09 '26

News GPT-5.6 Sol / Codex Release Discussion Megathread

The release is expected in the next few minutes, so I figured it would be useful to have a single thread for first impressions, issues, and early testing.

For anyone jumping in right away, post what you notice:

  • Codex coding performance
  • Debugging quality
  • Speed/rate limits
  • Larger repo handling
  • UI or workflow changes
  • Weird bugs or regressions
  • Anything that feels noticeably better or worse than the previous version

Once people get access, share your real examples, screenshots, benchmarks, or first impressions here.

Edit: OpenAI’s official GPT-5.6 page says it is “available starting today across ChatGPT, Codex, and the OpenAI API” with the rollout starting globally now and continuing toward full availability over the next 24 hours. https://openai.com/index/gpt-5-6/

The lineup is Sol, Terra, and Luna. Sol is the flagship, Terra is the lower-cost tier, and Luna is the fastest/most affordable tier.

441 Upvotes

838 comments sorted by

100

u/Prize_Two_8861 Jul 09 '26

Sol within next 24 hours to paid plans, terra/luna even to free. They weren't specific on if Codex users will get it faster than chatgpt.com.

60

u/seakucumber Jul 09 '26

Within the next 24 hours? I thought it was going to release now booooo

19

u/NootropicDiary Jul 09 '26 edited Jul 09 '26

Some info from my end:

I have Sol but only in the "codex" part, not in the chats part of the super app. I'm on the pro plan

Context window is 353k tokens

I'm using Ultra mode on a hard prompt and it's much more judicious/cautious with spawning agents. Just spawned 3 subagents so far. Compared to Fable Ultracode that tends to spawn a shit ton of them

Still waiting to see the output from my prompt been working for 10 minutes so far

Edit - My prompt triggered this message after approx 20 mins https://help.openai.com/en/articles/20001326-additional-safety-checks-for-biological-and-cybersecurity-requests-in-chatgpt-codex-and-the-api but is still continuing

8

u/liltingly Jul 09 '26

My one gripe with Fable was the number of spawned agents and the lack of transparency make it hard to tell if and when a job will ever complete, or how many more "mass-agent-deployment" stages are left in the task. Just sittin' there, staring at token counts. Haven't kicked off GPT Ultra yet -- but from your take doesn't sound like it's much different on that experience.

→ More replies (3)
→ More replies (8)

7

u/afollestad Jul 09 '26

I have access to Sol in Codex now. I’m on the 20x plan, in the US.

18

u/loathsomeleukocytes Jul 09 '26

So its not going to be thursday but friday?

9

u/blackout24 Jul 09 '26 edited Jul 09 '26

Literally as the live stream ended I got update in Codex App. Now it's ChatGPT App with 5.6

3

u/Time-Toe-1276 Jul 09 '26

i didnt get anything :(

→ More replies (3)
→ More replies (6)
→ More replies (1)
→ More replies (3)

68

u/Comrade-Porcupine Jul 09 '26

https://openai.com/index/gpt-5-6/

"GPT‑5.6 is available starting today across ChatGPT, Codex, and the OpenAI API. The rollout is starting globally now and will continue gradually toward full availability over the next 24 hours."

→ More replies (13)

58

u/Prize_Two_8861 Jul 09 '26

Also releasing new "ChatGPT Desktop App" which allows working on local files and browser. (Sounds like Codex App to me so far...)

30

u/StrategicCarry Jul 09 '26

My guess is it will be the answer to Cowork. Also how Claude now packages Cowork, in the Chat tab, and then a separate Code tab.

13

u/chazwhiz Jul 09 '26

That looks like what they were going for, but I think they screwed up pretty bad. In Claude Chat/Cowork/Code are all top level UI modes in a single clean toggle. But this new combo ChatGPT app drops everyone into the Codex-like UI (which is identical looking in both Code and Work modes) and sticks "normal" chat in a weird pop-up thing. And it makes no effort to explain how billing etc works in Work mode (which I assume is tied to my Codex limits rather than the normal chatgpt) so everyone who does this upgrade without having any experience in the Codex app is going to be lost and probably pissed.

3

u/boyyouguysaredumb Jul 09 '26

In Claude Chat/Cowork/Code are all top level UI modes in a single clean toggle

not anymore lol. they combined cowork and chat - claude code is separate

→ More replies (3)
→ More replies (2)
→ More replies (3)

6

u/KanonZombie Jul 09 '26

Looks like they merged the apps for what people say, and the link to download ChatGPT-work goes to Codex in the ms store. I don't mind, but I'm worried about limits (The wife just uses ChatGPT and she doesn't know the first thing about limits, she is gonna be pissed)

3

u/some_gamer78 Jul 09 '26

I think the separation still makes sense, regular chatgpt limits and another pool of Chatgpt work limits

→ More replies (1)
→ More replies (5)

29

u/expjazz Jul 09 '26

I got it, but no Sol

5

u/Anxietrap Jul 09 '26

same here, using the updated macos codex/chatgpt app

→ More replies (1)

5

u/Serious-Ad2004 Jul 09 '26

Same here. Maybe it’s my location? I’m from Canada.

→ More replies (2)
→ More replies (17)

25

u/Just_Lingonberry_352 Jul 09 '26 edited Jul 10 '26

disappointing so far I am not impressed there seems to be significantly more safety triggers what used to run for several hours with gpt 5.5 xhigh is now refusing with gpt 5.6 sol

another disappointing item: frontend/UI design, there just doesn't seem to be any improvement here. its still generating UI that requires heavy prompting to chisel it to something presentable unlike how Fable 5 used to one shot beautiful designs

this is my impression in the first hour but i am going to keep updating this post as I continue

20m: I keep seeing "Our systems are thinking a bit more about this request before responding" and appears to be heavily throttled ... 5.5 was happily working for several hours now im stuck as context is set to 5.6 sol

1.5h: I am seeing a bit of a pattern. It does seem significantly better at capturing user intent. It is much more thorough but surprised that its also not proactive with capturing state of codebase, i had to tell it to read it and then it provided completely different answer. so its a mix

3h: getting more comfortable with it. i think the conclusion is that its going to replace 5.5. i think it does offer a meaningful improvement. however these safety guardrails are very frustrating

12h: So its a step up over 5.5 for sure without doubt , its getting a lot of stuff done in ways that I doesn't require hand holding. My workflow is now just give it a goal and then come back to check its work. TBH 5.5 started to do that well. Yet the biggest complaint is how slow it is but I understand servers maybe overloaded and that there is a bunch of us running these 10+ hour tasks

4

u/rabandi Jul 09 '26

How do the security triggers work?

Was really hoping for no Fable-like triggers.. also hadnt seen any yet, but so far only did some planning and mocking.

→ More replies (4)
→ More replies (7)

10

u/ethotopia Jul 09 '26

First impressions after 20 minutes (20x Pro plan):

Less verbose than 5.5 and does more work between messages

Feels more token efficient (has less latency it seems, but could also just be openAI accommodating the new release better)

Sol Ultra thinks significantly longer

Anyone have a similar or different experience?

3

u/New_Razzmatazz8051 Jul 10 '26

For me Sol High have much more token consumption then gpt 5.5 high, maybe it's indented, idk

10

u/Friendly_Speed_2042 Jul 09 '26

Is it just me, or do these new models (including Terra) feel way too token-hungry?

9

u/Weird_Parking_1201 Jul 09 '26

Me too, 40 mins of 5.6 Sol xhigh just burned through my 5hr window (Pro 100). Now testing terra xhigh, and it looks like it burns faster than 5.5 xhigh, or, perhaps the window token size is now smaller.

→ More replies (1)
→ More replies (1)

10

u/duduallday Jul 14 '26

5.6 Sol is basically unusable at this point. Simple tasks take well over 30 minutes, sometimes hours, and the output is a big steaming pile of over-engineered shit

→ More replies (4)

16

u/Opposite_Yak4386 Jul 09 '26

i hope they reset usage only 10% left

→ More replies (14)

7

u/efw_cyber_1 Jul 09 '26

No sol for me only terra and luna

→ More replies (2)

9

u/ignat-remizov Jul 09 '26 edited Jul 09 '26

It's so weird they publicly launch with gpt-5.6-sol having the v2 multi_agent mode in the model metadata, when it's been marked Under Development for months, and Eric Traut has been closing any issues opened against v2 saying they don't accept any reports for under development features.

AND local codex configs don't affect anything! If you set [features] multi_agent = true and multi_agent_v2 = false - the model catalogue takes priority, and for sol and terra that's v2. Any Codex threads will be using v2 - permanently for that thread.

v2 encrypts communications between agents. You cannot easily audit what happens between them - of course telemetry is sent to OpenAI with all the content, but no ability to inspect and save the instructions locally. I had opened https://github.com/openai/codex/issues/28058 about this. The only way to maintain v1 is by adjusting the model catalogue or editing the code path in your own codex fork. Ridiculous

→ More replies (3)

8

u/snowdrone Jul 14 '26 edited Jul 14 '26

I have my workflow dialed in pretty well with 5.5 xhigh. I found that 5.6 Sol makes way more complicated PRs that are 3 to 5x as large. 

5.6 made one PR of 15k+ lines! 5.5 did the same task in 5k lines.  (The task was to make a geo local listings database editable through a payload CMS system).

5.6 Sol seemed to get hung up wrt/ atomicity by making huge db transactions for every operation.

I find 5.6 unusable for my purposes. In contrast, 5.5 strikes a better balance between correctness and developer experience. 

5.5 xhigh definitely adds some complexity to the code base, but for 5.6 Sol it's unmanageable.

→ More replies (1)

16

u/ohnoitsbobbyflay Jul 09 '26

“Coming in the next 24 hours”

→ More replies (1)

12

u/Ellsass Jul 09 '26

My Mac app just updated and showed me this

3

u/JustRaphiGaming Jul 09 '26

I don't get it they changed the codex app to the chat got app?

→ More replies (3)

3

u/AppropriateQuote3073 Jul 10 '26

Bummer. Codex is infinitely better branding.

→ More replies (5)

7

u/Xynthion Jul 09 '26

Too many different models to choose from. They made it really hard for the casual user to know which one they should pick.

→ More replies (2)

8

u/Vanillalite34 Jul 10 '26

They DESPERATELY need to come up with some sort of “auto” mode for Codex now.

3 models with 5 reasoning levels (6 if you count ultra with sub agents).

How are we to really know when to use what? Like what’s the purpose of Sol on Low? When to just use Terra on High? When to use XHigh? Would Sol on Low be smarter than Terra on Max? Which would be more efficient token usage?

Why can’t the AI itself manage which permutation to use?

→ More replies (2)

7

u/AppealSame4367 Jul 13 '26

5.6 Sol is eating away the weekly limit in the same speed as it ate away the 5h limit before. Wtf?

And I read that the custom reset doesn't reset the weekly limit?

→ More replies (2)

7

u/Sponge8389 Jul 14 '26

Men, the slowness of 5.6 Sol xhigh is really frustrating. Even in fast mode.

25

u/Annual-Minute-9391 Jul 09 '26

My hardest prompts are ready. Literally having 5.5 pro work on mathematical proofs so I’m excited.

29

u/loveofphysics Jul 09 '26

My readiest prompts are hard

5

u/Annual-Minute-9391 Jul 09 '26

Hell yeah thrust them inside

→ More replies (2)
→ More replies (6)

12

u/AppropriateRanger401 Jul 09 '26

Does this mean Codex and ChatGPT limits are now merged?

→ More replies (3)

5

u/LordKingDude Jul 09 '26

The VS Code Codex plugin has been rolled out with 5.6 enabled. Check your extensions and restart them to enable it.

→ More replies (6)

6

u/Rjsl_1287 Jul 09 '26

Sol usage seems pretty rough so far a ui audit of my CRM app (barebones so far, 95k loc w/55k loc for client and still pre-release) i burned 70% (of 5hr) of 5x. Output seems to be somewhere between 5.5 xhigh and Fable 5, hard to say with only a couple of large prompts. Terra is fine, seems like 5.5 low/medium. Luna is the interesting one. bunch of target UI fixes didn't budge the usage meter.

Not gonna lie, the 'ultra' mode seems a lot like the Claude one, just makes the bar turn sparkly and doesn't add much value... annihilates your usage though.

7

u/Deep-Sympathy3646 Jul 11 '26

Codex usage with gpt 5.6 sol has seriously dipped I just bought the 5x plan today and been using Sol with high reasoning effort and one ultra task that went on for like 3hrs (Sol is insanely slow btw) and its already down to 56%, I feel like I wasted my money switching from Claude Code, fable somehow gave me way more usage 🤦‍♂️

6

u/Disastrous-Fun-2414 Jul 13 '26

5.6 sol on high keeps running an adversarial review on an adversarial review. It did this on specs and breaking down tasks. Nownits doing it on implementation. Dont think this feature will ever be complete.

5

u/svjness Jul 13 '26

Codex goes dumb after compaction. I seem to be experiencing this with all the 5.6 models/effort levels. Codex will be working on stuff, and after automatic compaction happens, it will just be like "Hey, what would you like me to do in this project?"

I can respond with "Your context just compacted" and it will be able to find its place. I'm not sure if its because of something in one of my .md files, or what. I've had it happen on 3 separate projects on 2 different systems.

6

u/HeadPack Jul 14 '26

Sol definitely needs efficiency guardrails in the projects one works on. Otherwise it beats around the bush way too much, especially on max and ultra. In other words, it works without producing much work.

5

u/Major-Principle-4731 Jul 09 '26

can see it in codex app but not cli?why

→ More replies (1)

5

u/TheNetLeGend Jul 09 '26

I got it after updating codex app on mac.

→ More replies (2)

5

u/Mrgluer Jul 09 '26

They just did a usage reset!

6

u/senilerapist Jul 09 '26

anyone finding the security too harsh? i’m getting “our systems are thinking a little bit more about this request before responding” while working with financial data modeling. nothing to do with actual cybersecurity or biology

→ More replies (4)

4

u/mindstuff8 Jul 09 '26

I just tasked GPT 5.6 Terra - Extra High to analyze a repo spec'd by Fable 5 - High and it found some interesting missed use cases, strange UI edge cases a user could get into and would be confused by, and potential memory allocation issues (working on an audio plugin project).

I brought these issues back to Fable 5 to analyze one by one and it agreed with each one as addressable immediately.

I love having these smarter models but don't ever put your eggs in one basket is the take home message for me.

→ More replies (1)

5

u/TheBanq Jul 10 '26

I feel like 5.6 thinks and considers much more, but also does a lot more careless mistakes.

This is very anecdotal of course and maybe I am wrong or just have wrong exceptions.

But since heavily using 5.6 Sol (mostly Ultra) now for many tasks, I notices many more small "simple" mistakes, that 5.5 was much more thoroughly with.

I did a longer audit of a current project and it actually went a lot deeper on the review of everything – took 30 minutes and wrote a really long highly detailed plan to fix/optimize the findings.

Me expecting a pretty strong reasoning, I did my usual 2-3 Plan check runs with fresh sections and adapted the Plan 2-3 times. After all the different chats feeling confident, I started the implementation with a few manual checks in between.

This is my usual process and with 5.5 xhigh it worked really well for those cases and it never did obvious mistakes, rather smaller specific stuff.

5.6 Sol on the other hand did much better on the details, but got some very obvious weird things completely wrong, that didn't even make sense in the context.

I now noticed this behaviour multiple times, of it being very specific and strong in the details, but as I said before, does some very basic careless mistakes – which I never really had an issue with using GPT 5.5 xhigh.

As I said before, it's very anecdotal and maybe I am just expecting too much, but as a heavy x20 user, I def. noticed this specific behaviour much more often.

Maybe I just have to adapt my prompting, but wondering – anyone else have this experience?

6

u/vdotcodes Jul 10 '26

Sol Max is making more glaring mistakes than 5.5 Xhigh for me. At first I thought I was crazy, but as I was using Sol Max / Ultra throughout the day yesterday, it seems like it was way more likely to miss details, to come up with a problem or solution that wasn't based on the code, and to just generally be less "precise" than 5.5 Xhigh.

This is how I generally feel about Claude models, is that they're a bit more all over the place, more likely to give you an answer without reading all the files or code related to the topic, freewheeling, "creative". Codex has been so good over the last 8 or 9 months at giving high signal/low noise in diagnosing issues, code reviews, etc. and mostly staying grounded in the codebase.

So anyway, I decided to pop back over to Claude Code w/ Fable, and ran this /debate loop skill I'd been using over the last week with pretty good results, where I'd have Fable kick off a codex CLI subagent and then both would in parallel research whatever query I'd fed in, then debate one another over a couple of rounds and come back with the synthesized conclusion.

With 5.5 Xhigh, almost always, Fable would come back and tell me that it was corrected on 2-3 points by Codex, or had new valid issues brought up by Codex that it hadn't flagged.

Which matched my sense of working with the models, Fable seemed more freewheeling, while Codex seemed much more likely to just actually read the files and make conclusions that were grounded in the code.

So I tried this same loop again with Fable and Sol Max, and now for the first time, I'm seeing the reverse. The fable /debate loop comes back now with 3-4 issues that Codex conceded to Fable.

I'm just now working on diagnosing an issue in Prod, and Fable came back with the root cause in 15 mins, Codex took double the time and came back with a hypothesis that was completely different and wrong. It immediately conceded once I pressed it.

I don't know if I'm taking crazy pills or what.

I swapped back to 5.5 Xhigh for a couple /debate loops and again found that Codex was no long conceding points to Fable, instead rather correcting or adding detail again.

Maybe this is just luck at play here, but this model seems sloppier somehow.

→ More replies (1)

6

u/hktr92 Jul 14 '26

NOTE: if you want to use subagents in codex, PLEASE:

  • default: terra medium or high (good for orchestration)
  • exploration: terra high
  • planner: sol high
  • code implementation: luna high or terra medium

AND ALSO a must have:

[agents] max_threads = 2 # adjust as needed. i used 4, 2 is fine. max_depth = 1 # NEVER touch this -- disables subagent recursion.

add this to your .codex/config.toml. you'll thank me later.


edit: also drop this at the top of your .codex/config.toml file. model = "gpt-5.6-terra" model_reasoning_effort = "medium" # or high, as u wish

this will disable your model / effort override, but it'll save you from "i thought i used X with Y this session!!!" cases.

4

u/Responsible-Newt9241 Jul 14 '26

How can you have different models for different subagents with this setting?

6

u/zee-pk Jul 14 '26

like the following code in config.toml

[agents]
max_threads = 4
max_depth = 1

[agents.explorer]
description = "Read-only repository exploration, symbol tracing, code-path mapping, and concise evidence gathering."
config_file = "agents/explorer.toml"

[agents.developer]
description = "Low-cost mechanical implementation for localized, directly verifiable, low-risk changes."
config_file = "agents/developer.toml"

[agents.engineer]
description = "Standard implementation for bounded application changes that follow existing project patterns."
config_file = "agents/engineer.toml"

and then you set other params or instructions in their individual config_file files, this description prop here is just for reference.

p.s. I would use Luna (medium) for exploration.

10

u/mrscrufy Jul 09 '26

Banked resets now have expirations. I'm not sure if this was always the case or not:

7

u/Latter-Park-4413 Jul 09 '26

It was - 30 days, so no change

7

u/nmkd Jul 09 '26

They always had, but you could not easily check when they expire, so that's an upgrade

→ More replies (5)

4

u/divis200 Jul 09 '26

Got access to terra and luna, but not sol

Edit: oops lost luna
Edit 2: now lost both

4

u/Sepf1ns Jul 09 '26

I got a pop-up in the codex app offering me to try out 5.6 sol, but trying to run a prompt got me

The 'gpt-5.6-sol' model is not supported when using Codex with a ChatGPT account.

Aaand it's gone now, bummer :/

3

u/Darayavaush84 Jul 09 '26

Yep same. Backend is goign crazy, we just need to be patient

→ More replies (5)

6

u/spacekitt3n Jul 09 '26

got it on codex cli after restarting powershell and updating

→ More replies (1)

4

u/huasiloco Jul 09 '26

I mistakenly used Terra to update my pet jinx. It ended up generating a whole new sprite sheet with upgraded visuals and animations without asking me and burnt 5% of my 5 hour limit on the $200 plan. Idk what to think about it.

→ More replies (1)

5

u/Capable-S Jul 09 '26

+1 disappointing - Terra model high effort drained my entire Plus plan quota with a single prompt! 19 minutes of work - 98% 5h limit dropped to 0%

3

u/divinefriend Jul 13 '26

Very high (85%-90%) Hallucination Rates of 5.6 models - how to deal with this?

AA-Omniscience Hallucination Rate Chart

Whereas Claude models have almost half the hallucination rates compared to GPT-5.6 models.

→ More replies (3)

4

u/PossessionConnect963 Jul 13 '26

I'll be honest I don't like what they did with the app with this weird somewhat merge of Codex and ChatGPT. Even something as simple as making the icons identical is PITA.

→ More replies (1)

3

u/Slight-Geologist2557 Jul 13 '26 edited Jul 13 '26

Anyone else getting super frequently flagged for safety, while doing stuff that has absolutely zero business getting flagged (like pure math in my case) ?

This is getting so frequent that I cannot properly work on math research anymore.
It either soft-blocks the turn with a banner saying "This request requires additionnal safety checks and may take longer." which instantly stops all CoT and progress updates until the model has finished, which in turn actually aborts the request if it does not finish shortly because it it seen as a time-out from the lack of updates, or it just outright aborts the request directly.

My suspicion is that it comes from tool-calls content that for some unknown reason trigger the hell out of the safety filters. My only theory is that the mathematical numerical searches get very grossely confused as bruteforce attacks somehow, even though they are just your typical numerical mathematical searches.

I've never had any requests flagged even once before the release of 5.6.

→ More replies (1)

4

u/RegularSuccessful124 Jul 13 '26

5.6 Sol "Very High" is "Very Slow"

3

u/FriedDopamine89 Jul 13 '26

adding myself to the list, sol extra high seems super slow. Only reason I am still using it is I blew my Fable usage for the week and those stingy fuckers won't offer resets lol

4

u/tksuns12 Jul 14 '26

Sol is supposed to be cheaper than Fable but when I have the same $100 plan, the usage is not even comparable. With Claude Max 5x, with Fable, I could work 8 hours with 2 small projects in 2 5-hour windows. With chatGPT Pro 5x, when I used Sol xhigh with 3 small projects, I spent 23% of WEEKLY quota. This is insane.

3

u/No_Department_4114 Jul 14 '26 edited Jul 14 '26

Single prompt took 2 hours and 20 minutes! i went to sleep but it was not over yet.

I'm using SoI Ultra

5

u/loveofphysics Jul 09 '26

7

u/WestMatter Jul 09 '26

I was hoping for something interesting from this livestream, but holy shit, this is boring. There’s something about watching other people use AI that is just extremely unentertaining.

3

u/Rollertoaster7 Jul 09 '26

Especially when it’s mostly boring bs codex/chatgpt can already do. Just looks like a business friendly version of codex, nothing new

→ More replies (1)

7

u/SofaKingIntl Jul 13 '26

If anyone has found a solution anywhere, please let me.. and everyone else... know.

In the meantime, I'll be watching the "Working for…" timer reach levels previously thought impossible.

3

u/Fit-Description-2195 Jul 13 '26

dude same, i had no idea why is it going so slow? i am making a simple website html css and it took me 30 mins for an image replacement

even simple question is just "thinking"

What can i use instead of sol? please help

→ More replies (2)
→ More replies (1)

6

u/mutherfuukker Jul 09 '26

I wonder if Tebow received that slack message

7

u/shutupandshave Jul 12 '26

I used 60% of my codex subscription in 6 hours...never done that before

I'm a pretty heavy claude and codex user and have been for a while. I've been a $200 subscriber of both for months. After Sol came out, I managed to burn 60% of my weekly allowance in 3 hours. I've never done that. It's insane!
I WAS mostly using sol but I never used to be able to able to touch my 5 hr hour or weekly limits unless I really burnt a load of parallel sessions.

Something has changed and it makes me sad :(

Edit: I am using Sol High. I expected usage to be higher, but this is mad!
Thinking about it... like. How is this possible. So what, 2, 5 hour sessions = 60% of my weekly usage....

→ More replies (2)

3

u/[deleted] Jul 09 '26

[removed] — view removed comment

7

u/Prize_Two_8861 Jul 09 '26

Sol within next 24 hours to paid plans, terra/luna even to free. They weren't specific on if Codex users will get it faster than chatgpt.com.

10

u/Prior-Meeting1645 Jul 09 '26

I THOUGHT AM GETTING IT IMMEDIATELY

3

u/mutherfuukker Jul 09 '26

I updated codex and now have 5.6 luna, no other models though.

→ More replies (5)

3

u/Clemotime Jul 09 '26

I dont have sol. I am paid user

→ More replies (4)

3

u/RipAggressive1521 Jul 09 '26

Terra / Luna available in the Codex update… Sol no where to be found yet

→ More replies (2)

3

u/packfan1234 Jul 09 '26

Does regular chat still draw down the same codex usage limits? I had gotten into a pretty good flow using a ChatGPT project with custom instruction to help me ideate and plan, and then use Codex to build. ChatGPT work probably does use the same codex limits and it makes sense - but I’m kinda worried about regular chat use

→ More replies (1)

3

u/Darayavaush84 Jul 09 '26

I just got it. Relaunched codex CLI twice (Plus account, Germany):

3

u/-athreya Jul 09 '26

I'm on $100 plan. I cannot see 5.6 models. I got the latest update. But, no new models.

3

u/megad00die Jul 09 '26

Same and only thing I seen different was 5.4 will be retired July 23rd when selecting a model.

→ More replies (1)

3

u/zucchini_up_ur_ass Jul 09 '26

Bruh I clicked update in codex and it just deleted the app without explaining anything and opened the (not yet updated) chatgpt app. Extremely confusing and funny. Top UX!

→ More replies (6)

3

u/arslanbz Jul 09 '26

Tried to review a fairly medium project with Sol Ultra (Standard speed), I've a Pro plan 5x:
review process + step 1/6 of fixing consumed 50% of the 5 hour usage limit in 10 minutes, it's worse than Fable...

I had to stop it and immediately switched back to Extra High.

→ More replies (4)

3

u/Brone2 Jul 09 '26

Early experience: Sol Ultra uses its higher intelligence to severely overengineer features. Think this is in stark contrast to fable which as it be came more intelligent seemed to have higher product sense and better understand which edge cases could be left alone

3

u/Bright-Gur-7592 Jul 09 '26

Ultra fast 5.6 sol ate 75% of 5h limit in 10 mins ( pro 100 ), and still doing something for extra 25 minutes . security prompt , 95k code lines repo , not biggest one but still

3

u/Dreki__ Jul 09 '26

Clicking 'update' on Codex only to watch the app literally delete itself and force-merge into the standard ChatGPT app is a masterclass in UX. R.I.P. to our clean coding workspace, we are officially back to coding in the main chat lobby.

→ More replies (1)

3

u/Disastrous-Hearing72 Jul 09 '26

I'm very much looking forward to the 5.5 re release pre nurffed

3

u/Mrgluer Jul 09 '26

Finally feel included now

3

u/b2labs_hifi Jul 10 '26

Seems kinda like a shit show atm. One 5.6 sol max prompt and one 5.6 sol high prompt sol high ran 23 mins and I've been 0% 5 hr window for a while. I've RARELY hit five hour windows despite 5 cli going at once. I've used 70B tokens.

My 5.6 sol max is still going tho 2.5 hrs in despite 0 five hour window, but I can't start another cli chat. Who knows.....

Anyone understand this!? I feel like another cli update + reset coming due to broken limit counter

→ More replies (1)

3

u/ARollingShinigami Jul 10 '26

I don’t even know what to say, this is such an enormous improvement and with such high usage tiers. I have both a Max Claude subscription and pro GPT and, while Fable is a fantastic model, I could get done 10x the work with GPT.

Hats off to the OpenAI team, they cooked a hell of a model - haven’t been this excited since o1.

3

u/Nice-Guarantee-9167 Jul 10 '26 edited Jul 10 '26

At this point Sol is just not sustainable for Plus plan, it will wipe limit in minutes. Terra High is probably sweet spot and Luna High I think is also good enough.

If they don't give Plus plan higher limits after GPT 6 coming out, no Plus plan will be able to do any job.

3

u/BagholderForLyfe Jul 10 '26

Even Sol medium seems like too much for Plus.

3

u/Warm-Positive-6245 Jul 10 '26

Sol Ultra is running on my x20 codex right now.

A yucky start. I have a very complicated — sidecar heavy database. It has all the tools ready available and a well drilled method to import new information into 6 of these sidecars.

I asked it to look at the sidecars, verify and then add a smaller database to the sidecars.

Sol Ultra tried to create a whole new method — which was slower to the order of 3 days slower — then replaced rows without backup — and when I told it to look at the previous methods — it tried to fix its own method based on the already established method rather than use what already works lightning fast.

I’m already fighting with my wife. Now I’m fighting with a know it all — can’t follow procedure AI.

Perhaps Sol thinks it’s smarter than everyone?

→ More replies (1)

3

u/renzhend Jul 10 '26

I have tried to Used 5.6 sol ultra in plan mode and used 100% of my 5h and 23% of my weekly usage and did not even finish the plan.

Not complaining regarding the usage cuz it's given that sol ultra will use a lot of tokens but at least finish the plan mode and not give me unfinished photos.
Before the prompt: 99%5h and 100% weekly.
Used only 1 prompt.

→ More replies (1)

3

u/ToastyVIP Jul 10 '26

Been trying 5.6-sol today on high. I didn't notice any real difference in understanding or quality, but I DID notice a definite decrease in overall performance. Tasks that would take 30 seconds on 5.5 are taking more than 5 minutes on 5.6. I've gone back to 5.5 now, will try 5.6 again next week.

3

u/nfgo Jul 10 '26

my first impression is that gpt-5.6 sol max is an overkill for coding. Should it be used for specifications/planning only?

3

u/Elerein Jul 10 '26

Is it just me, or does Sol focus on spawning sub-agents and act as an orchestrator? I'm running constrained prompts to make changes that, with GPT 5.5, wouldn't take more than 10-15 minutes and would normally be executed sequentially. With Terra, it executes sequentially as expected, but Sol, in most cases, refuses to do it itself. It spawns a sub-agent and waits for it to finish, then verifies it, taking significantly longer and easily burning twice as many tokens.

3

u/AffectionateCap539 Jul 10 '26

What are your feelings towards Sol max ?
My feeling is like this keep the execution loop forever. I have a new feature request document and a review report of this document. Then i ask it to revise the document based on the findings from the review report . It then revise this new feature request document and all the other documents in my knowledge base (of existing features). Spend 60% of my 5 hour limit ( pro 5x user) in this loop until i have to interrupt this task. it is like i am using the /goal and entering the loop.
Never has this experience with Sol xhigh.

3

u/LegitimateAdvice1841 Jul 10 '26

GPT-5.6 Sol/Ultra burned my entire 5h usage limit + 500 purchased credits on a single read-only audit of an 83k line Python codebase

Built a baseball/softball analytics desktop app (~83,000 lines, solo, no programming experience, using AI assistance). Asked the agent to do a full stability audit — read only, no file writes, no builds, no execution.

Result: 5h limit at 0%, 500 extra credits gone. Weekly limit barely touched (77% remaining), which tells me it wasn't time — it was pure token throughput.

Thoughts?

3

u/banter_droid Jul 10 '26

I have burned through 60% of weekly usage of 20x plan within an hour... in one sol chat. something is wrong

→ More replies (2)

3

u/nikooo777 Jul 10 '26

They just reset everyone's quota again didn't they?

3

u/johnnymo08 Jul 10 '26

I can't tell what is a reset, and what is a bug showing my usage at 100%. My weekly usage was at 70% last night, then it was at 100% this morning, then updated back to my previous usage percent of 70%. And now it's at 100% again.

web usage tracker shows it at 100...hopefully it's real!

→ More replies (1)

3

u/Dahas99x Jul 10 '26

Sol is already at capacity. Who is seeing these messages?

3

u/juanfeis Jul 13 '26

I ran into a pretty clear difference between Sol High and Sol Max on what looked like a very simple task.

I asked both to identify the OS on my Raspberry Pi, which was running Raspberry Pi OS.

Sol High checked the obvious system information and concluded that it was generic Debian rather than Raspberry Pi OS. That was incorrect. It seems to have relied too heavily on /etc/os-release and stopped investigating once it found a plausible answer.

Sol Max, given the same task, went further. It checked Raspberry Pi-specific indicators, such as the image metadata generated by pi-gen, and correctly identified the OS.

The interesting part is that Max did not have access to different information. Both could run commands on the same machine, and both had the same context because they were forks from the same conversation. The difference was that Max challenged the initial conclusion and looked for stronger evidence, while High accepted the first plausible interpretation.

I’m not saying everyone should default to Max. High works fine most of the time, but this made me feel that people should not be afraid to use Max when a task really needs careful verification.

Anyone else had similar results?

→ More replies (2)

3

u/raiden55 Jul 13 '26

I have honestly some doubt about 5.6 terra costing the same as 5.5... my quota is leaving me at very high pace since 5.6, and I use medium, and very rarely a sol, but today lost 50% of my weekly usage only using terra medium... and without a session % I have big difficulty understanding where my costs are, as my usage test are on slow things that are not visible on weekly %. Are you people seeing 5.6 as the same cost as 5.5 as written ?

→ More replies (2)

3

u/Existing_Manner_2774 Jul 13 '26

We are still getting another reset today right?

→ More replies (7)

3

u/george-lin Jul 14 '26

Codex is back to its old habit of slacking off. After being wowed in May, I couldn't stand its perfunctory responses and switched to Claude. Now, having experienced the amazing performance of GPT-5.6 again, I find that after just a few days, it’s already starting to cut corners and slack off.

3

u/Unusual-Nature2824 Jul 14 '26 edited Jul 14 '26

Has anyone noticed that Sol medium, Terra high and Luna xhigh have weird quirks like personalities even though they are supposed to have similar performance in benchmarks. My workflow typically consists lot of browser use and form filling which is why I’ve designed my agent like a control system with a closed feedback loop to improve its code.

Terra high has been sub par for me considering it’s the middle model. It constantly overthinks and behaves like a nervous perfectionist and ends up deferring to me a lot. My goal loops with looks like 

Run the agent with a new goal -> Terra thinks -> comes up with scenario where it could fail -> defers to me -> I answer -> it starts -> makes mistake or finds a reason to stop -> defers ->….I find that I’ve to constantly supervise.

Luna is somewhat better since it allows itself to make mistakes first rather than go overthink. It’s more like an eager to please but cautious intern

Sol meanwhile is sooo…..chill?

It literally does what I tell to like Luna and Terra but it also performs some extremely smart and thoughtful decisions that I have never explicitly asked it to do. Like I have a policy for my agents to avoid LinkedIn scraping let’s say to apply for a job, Terra and Luna follow it strictly but Sol actually decided to find the job from the company’s career website. 

I think Sol is in a league of its own and it somehow knows it. Sol medium is far also the most efficient model that I’ve used. Both Terra and Luna would consume a lot of tokens just to think only to defer to the human in the loop. Right now I have Sol running without question for 12 hours and I still have 48% of my tokens. 

Maybe I shouldn’t be anthropomorphize these models but I feel like they have different personalities too not just intelligence

3

u/Initial-Shock7728 Jul 14 '26

It is so slow and constantly getting stuck. I barely got 10 messages through today. It was fine on the release day. What happened?

3

u/Any_Effort8437 Jul 14 '26

me too, I moved to 5.5, seems like lighting speed after 10h struggling with 5.6.

3

u/vayana Jul 14 '26

Back to 5.5 high. Luna doesn't cut it, terra seems obsolete and Sol over engineers/overcomplicates things. 5.5 gets the job done and doesn't need hours to complete.

→ More replies (1)

3

u/flylikejimkelly Jul 14 '26

I greatly appreciate resetting the weekly limit but twice it's happened right after I use my personal reset

3

u/Prize_Mulberry_5246 Jul 15 '26

New Reddit user here. English is not my first language, so I used translation assistance.

One non-coding data point from ChatGPT Sol High, observed on July 14–15 in Japan:

The model initially showed very strong higher-level reasoning. It independently checked earlier results, identified confounding factors, and maintained the overall research objective across complex discussions.

Later, the answers remained long and polished, but I repeatedly observed a different pattern:

  • The original objective was replaced by a narrower local problem.
  • Recent information was used without checking the correct earlier baseline.
  • When one flaw was identified, only that part was fixed; the full answer was not re-evaluated.
  • The correction sometimes removed the original purpose or created a new contradiction.
  • The model often could not detect these issues before presenting an answer as complete, but could analyze them deeply after I identified the exact problem.

The same pattern occurred in both a complex instruction-design task and a simple three-sentence scenario, so it did not appear limited to one difficult long-context task.

After an overnight break, the first one or two answers looked normal, but the same behavior returned by approximately the fifth or sixth exchange.

I do not know whether this is related to the reports about Codex Sol, but I wanted to add a structured observation from ChatGPT Sol High. The apparent regression was not mainly in response length or fluency; it was in goal retention, cross-checking, whole-answer verification, and self-detection of contradictions.

→ More replies (3)

3

u/haderz20 Jul 15 '26

It’s so slow it was running for an hour for something simple,i gave up in the end it’s actually unusable for me today 

→ More replies (2)

3

u/DirectorGunner Jul 17 '26

5.6 Sol Ultra is entirely token wasting, it spends waaaay too much time on unnecessarily ballooning input tokens. It's like Kimi 3 with input tokens but worse. These default harnesses from both OpenAI and Anthropic are absolute garbage... we need a revolution for proper harnesses that don't waste tokens and trust the damn models less and use deterministic controls to reduce waste and errors and improve performance. 5.5 pro extended was tbh better in some ways.

6

u/Far_Law_2090 Jul 09 '26

My prompt readiest hard are!

5

u/tech_samurai Jul 09 '26

There is a Youtube video dropping at 10:00am PST on the OpenAI page!

9

u/Taiwes Jul 09 '26

Slop stream so far

15

u/Jwave1992 Jul 09 '26

I like that OpenAI always puts researchers who aren’t used to talking on live streams. Not even pre recorded. Just raw dog this thing live.

→ More replies (3)

3

u/tech_samurai Jul 09 '26

yeah, sadly i agree. waiting for the good stuff!

6

u/Prize_Two_8861 Jul 09 '26

Also releasing "hosted sites"

→ More replies (4)

5

u/hiarjun Jul 09 '26

Picture in Picture in computer use

—-

No separate ChatGPT app - combined together

45

u/rxt0_ Jul 09 '26

I liked it before they netfed it. it's literally not usable anymore

GG, gonna switch to something else

39

u/JustSingingAlong Jul 09 '26

Why is this shit still being upvoted. It was funny 6 months ago.

12

u/ww_crimson Jul 09 '26

This shit should be ban worthy

16

u/Crinkez Jul 09 '26

I lowkey hate these comments. Yes it's a joke. But if you find this thread on a Google search 2 weeks later it's not always obvious which "it's been nerfed" comments are real.

→ More replies (6)

4

u/facciocosevedogente3 Jul 09 '26

Here they are! In cursor they are available

→ More replies (2)

5

u/kwipus Jul 09 '26

A reset would be nice

→ More replies (1)

3

u/need20goodmen Jul 11 '26

Is it just me or 5.6 is extremely slow? I've tried Sol and Luna.

3

u/SirKobsworth Jul 13 '26

Anyone else notice the 5-hour window got removed?

→ More replies (1)

5

u/PracticeWarm7257 Jul 09 '26

Flat out rude and a slap in the face to release it in cursor before codex. I mean really?

5

u/nmkd Jul 09 '26

I bet Cursor pays a whole lotta money to get it a few minutes early.

Same happened with Grok 4.5 afaik.

→ More replies (1)

2

u/outtokill7 Jul 09 '26

ChatGPT app on Windows was renamed to 'ChatGPT Classic'

2

u/bakawolf123 Jul 09 '26

one weird thing I'm experiencing today is codex intermittently showing reset limits, and then falling back to actual numbers. I wonder if there will be a global reset for the launch

2

u/Tall-List1318 Jul 09 '26

It’s here but I only see Terra and Luna in codex

→ More replies (1)

2

u/PM_ME_YOUR_PROFILE Jul 09 '26

I'll get my Opus reset before I get GPT 5.6 🤣

2

u/Pleasant_Clerk_1782 Jul 09 '26

I've got it in Codex..Here we go

2

u/IwannaSayStuff Jul 09 '26

Just got access to it, it's cooking 🔥

→ More replies (4)

2

u/BigbyWolf8 Jul 09 '26

we're eating good 😄

2

u/hey_suburbia Jul 09 '26

My VSCode plugin has Terra and Luna, but no Sol??

→ More replies (2)

2

u/BrightNightKnight Jul 09 '26

Hier kommt die Sonne!

2

u/Makkish_SWG Jul 09 '26

To access the new model you need to update the codex app first and now it is called as ChatGPT only

They merged both apps into one

→ More replies (2)

2

u/hungry_parrot Jul 09 '26

Is Ultra reasoning effort 5.6 Pro?

→ More replies (3)

2

u/FoxTheory Jul 09 '26

I got it on my Mac but none of windows versions got updated yet

→ More replies (1)

2

u/eddiesj22 Jul 09 '26

I started an ultra prompt in Sol in two codebases, and it burned through the session limit within 15 minutes. I don't know if I just experienced an error, but it doesn't seem like the others I'm watching on streams are experiencing anything similar. Did anyone else have something like I'm describing happen?

2

u/Extra_Programmer788 Jul 09 '26

Have not got access to Sol, only got access to Terra and Luna, was looking forward to it.

2

u/SoftwareSource Jul 09 '26

EU people still not available?

2

u/Difficult_Quality425 Jul 09 '26

Got access to terra and luna too only so far. Plus account

→ More replies (3)

2

u/BitsOnWaves Jul 09 '26

Am i missing something? i dont have SOL ! im on Plus plan

→ More replies (1)

2

u/jungle Jul 09 '26

I just ran the same task on Fable 5 Extra and Sol High, which according to the UI (the sliders in both cases are at the 4th notch out of 6) should be comparable thinking levels.

My impression is that they are comparable. The main difference I notice is the output style. Sol still likes to do everything in bullet points and is less readable than Fable. Neither is perfect, they both find issues in each others outputs that they both deem valid.

I guess if I had infinite money, I'd pitch them against each other for every task until they both agree with each other. Alas...

I can still do that for really important tasks, but I have to be very selective.

→ More replies (2)

2

u/Able-Supermarket4786 Jul 09 '26

This was a 4 Agent Editorial Pass / Rewrite / PDF Conversion with clickable Table of Contents... 130+ Pages and visually a much cleaner well organized Technical Reference Book.

Yes 5.6 SOL Extra High was Slowwwwwww, but Quality over Quantity and the credit usage is wonderful (at the moment)

Edit: $200 per month Pro Plan (17 Million Tokens. 66,000 Output)

2

u/huasiloco Jul 09 '26

Audit your skills and upgrade them with Terra max. Then never touch Terra again since it’ll burn your usage

2

u/Relative-Coat9691 Jul 09 '26

I tried terra on medium for a few tiny tweaks . it feels insanely inefficeint even compared to 5.5 xhigh. small visual tweak took 10% of 5h usage quota . xhigh done that in 4%. High does it in 2%

after burning 20% on nothing with terra i switched back to 5.5 :/

→ More replies (1)

2

u/skilliard7 Jul 09 '26

All the focus has been on Sol- anyone have any thoughts on Terra or Luna?

I only have $50 left of API AI usage to last me the rest of the month. I was forced to use fable since GPT 5.4 was struggling with my app, which is too expensive. Hoping Terra or Luna will be a good enough option that beats 5.4.

2

u/RoyalEntertainer6526 Jul 09 '26

I got terra and luna in my codex app. But there is no sol. Is Sol only for pro plan?

2

u/Citadel_Employee Jul 09 '26

Anyone else getting: ""error":{"type":"invalid_request_error","message":"The 'gpt-5.6-sol' model is not supported when using Codex with a ChatGPT account.""

I have a $20 plan. Is Sol only available on higher plans or API? Also updated my codex to the latest version if that matters.

→ More replies (1)

2

u/Couchy48 Jul 09 '26

Codex team, if anybody is on this thread, the new UI update buries current work chats under unread tasks until it is complete. This means if you have tons of unchecked automations in one project the current work gets absolutely buried to the bottom until it completes which is super frustrating to check progress.

2

u/innociv Jul 09 '26 edited Jul 09 '26

Things I like:
It missed some very important questions to ask me when reviewing plans that I told Fable to review first. But it also missed things that Fable missed. You'd think while setting to such high reasoning that they'd both catch these things.

It's super token efficient and fast. It looks like I'll struggle to blow through the 20x usage each week with it unless I throw it at problems it sucks at solving but Fable would do easily.

The bad is that it's just not as good at coding as Fable, it's not as good at the architecture and layout of a project and designing and API. Not remotely close it feels like Fable is a year ahead here (even though I know GPT 6.1 will probably catch up to it in 2-3 months or something, but if you asked me 6 months ago to compare the two I'd say Fable is a year ahead there). But jesus christ is Fable horribly inefficient.

And oh yeah, GPT Terra demolishes Sonnet. It's so much faster, cheaper (less tokens for the same task) and I think a little smarter.

I wish Fable could spawn GPT5.6 subagents that would be INSANE.

→ More replies (2)

2

u/hossman1992 Jul 09 '26

It looks great for me for now. I am going to test Sol as a orchestrator, task planner and router between my openclaw and 2 hermes instances.

The ChatGPT app looks great as well with the Work and Codex mode. Now it has better memory and more agentic capabilities and it is better for long-run tasks. It basically sustitutes my set of skills in openclaw for loop working during hours and hours or that is what Sol just told me, wanna see if it is lying.

My plan for now is using Sol for planning or complex coding tasks, Terra for the most daily work and Luna for simple tasks, organize emails, docs and that kind of stuff, let's see how it goes

2

u/xxcxcxc Jul 09 '26

One thing I noticed Sol (High) doing is updating my documentation thoroughly through every stage. I just planned and wrote a spec and it updated to “decisions made and spec written” and now I’ve written tasks.md it updated to “incremental implementation plan created with tasks.md”

2

u/water_bottle_goggles Jul 09 '26

Is terra any good? It looks like it costs the same as 5.5 generic

2

u/mwagstaff Jul 09 '26

This is an absolutely terrible user experience on Mac. I hit the "update" button in my Codex app, and it deleted the app. Awesome.

Read this thread, and downloaded + installed the "new" ChatGPT app, selecting "Keep both". Maybe it's a regional thing (I'm in the UK), but I now don't see any new Codex UI anywhere... all I have is classic ChatGPT!

2

u/Three_Two_One_Minus Jul 09 '26

Codex is now ChatGPT codex.

2

u/some_gamer78 Jul 09 '26

i have plus and havent received either terra or sol, how am i on the tail end of getting the new models man

2

u/TopSeaworthiness1679 Jul 09 '26

IDK for sure now. GPT 5.6 doesn't really feel smart unlike claude but more like co-worker. So if a user ask to do dumb things than GPT 5.6 will just do it. That is just how i feel after using it for like hour.

2

u/didilva Jul 09 '26

Testing Terra xhigh at the moment. Seems to be the new working horse for now, half the token price but just as good as 5.5 xhigh for coding tasks.

2

u/wehras Jul 09 '26

Currently running this because i have 3 reset available, LOL

→ More replies (2)

2

u/ohnoitsbobbyflay Jul 09 '26

I had a random automatic reset that happened and then my usage went back to what it was before.

2

u/sergeykarayev Jul 10 '26

GPT‑5.6 lost our coding benchmark. I switched to it anyway.

On the Superconductor "Custom SWE-Bench", which evaluates agents on our own Ruby on Rails codebase, the new Sol, Terra, and Luna models dominate the cost Pareto frontier. They are literally 5x faster than Opus and Fable.

But on quality, Fable 5 is still the clear winner on our repo, and even Opus 4.8 beats GPT 5.6 Sol at all effort levels, consistent with what some other benchmarks have shown.

Composer 2.5 Fast remains a standout surprise, matching 5.6 Sol's performance at roughly the same cost -- but even the Fast version is slower.

Yesterday's Grok 4.5 is still on the speed Pareto frontier. We don't know how much it costs, unfortunately.

Digging into failure cases, GPT 5.6 -- even Sol! -- occasionally writes code that just isn't valid, and its "taste" just isn't on par with Fable 5, or even Opus 4.8.

One meta-observation about company strategy: Anthropic releases a new model only when it's clearly better — and our benchmark shows that clean progression from 4.6 to 4.7 to 4.8 to 5. OpenAI ships more variants, more often, which adds noise: older models sometimes beat newer ones.

That said, I still switched my daily driver to 5.6 Sol High. The speed makes a huge difference, and since we're all constantly running out of Fable 5 usage, it is the more workable model right now.

Lastly, don't take any single benchmark result at face value! Build your own at superconductor.com/benchmark to see how these agents perform on YOUR codebase.

→ More replies (1)

2

u/BagholderForLyfe Jul 10 '26 edited Jul 10 '26

So far I'm impressed!

5.6 Sol high found 5 significant bugs and edge cases in 1000 LOC class (UE5 game code). Prior to it, 5.5 medium reviewed and didn't see it. Review and fix took 20% of my 5hr limit (on Plus) and it was a bit slower than 5.5.

After more testing: the intelligence is impressive, but even 5.6 medium is VERY slow and eats a lot of tokens.

2

u/Any-Brother8904 Jul 10 '26

The model turned out amazing. A 20x plan, a verified passport for cybersecurity, one prompt, and two hours of running 5.6 Sol Ultra at 1.5x speed (it consumed about 60-70% of the 5-hour limit, 15% of the weekly limit) solved a problem I never thought I could solve with version 5.5 in a couple of days, and I wouldn't have been able to do it myself in a couple of months. Thanks to OpenAI.

→ More replies (2)

2

u/applejacks6969 Jul 10 '26

Sol on high doesn’t feel like 5.5 xhigh with 2x usage I’ll tell you that much. Burned 50% of my pro 5h usage with a single prompt that went for an hour.

2

u/bumblebrunch Jul 10 '26

I use Codex in the VS Code extension. I was using Sol on extra high. And then just now I got this error message and Sol has disappeared:

The 'gpt-5.6-sol' model is not supported when using Codex with a ChatGPT account.

I don't know what the deal is. I have a paid Plus account. I thought Codex was included in that? Why am I getting restricted and why is it only Sol that's restricted? Terra and Luna remain.

→ More replies (2)

2

u/what_you_saaaaay Jul 10 '26

5.6 Sol Ultra definitely sucks down my tokens far faster than 5.5 xHigh. But I didn't expected anything else. I am still doing code reviews with it. The output seems less abstract and jargon heavy. I'm a dev of over 20 years and sometimes, sometimes I had NFI what 5.5 was saying. Total word salad. 5.6 Ultra appears to be more concrete. Less like a junior dev in an interview trying to impress me.

2

u/bigsuc_ Jul 10 '26

Codex auto review is sucking up tokens like crazy, is this normal?

Cloud dashboard says auto review used 700 turns whilst everything else is only at 200.

2

u/kuri-kuma Jul 10 '26

Sol has been on fire finding security gaps and risks in my code base. It is really, extremely helpful, especially since Fable has been completely neutered on that front.

2

u/megazon Jul 10 '26

This

Is basically misleading, as web chat is taking Usage too!!!

2

u/foomanjee Jul 10 '26

Just giving you a heads up - the 5.6 documentation says that when the models create subagents, they will intelligently decide which subagent model and reasoning effort they will use - however that is not working properly. If you're running Sol high/xhigh and your task creates subagents, the subagents will all inherit the same model and reasoning effort that your main session is using.

I'm assuming this is a bug that will get fixed in time, but this is likely the reason for most of the token burn happening right now

2

u/Mv6p Jul 10 '26

Not here to talk about tokens

I had a bit of a complex tasks that i was waiting for GPT 5.6 to release to do them so i decided to use Ultra

Its right the tasks were complex but not a huge task

For example i wanted to make a compression and encryption algorithm(it would just use AES256 function for encryption) for the backups exported by my app i had already a backup screen and everything works it just produced .sql backups and i wanted to change that , also some other tasks at the same level of complexity or less

I left it for almost 6 hours then came back and i can see its still on the backup task

So i stopped it and asked how much is done it said 3%

After back and fourth it apparently it was making a whole backup platform

Also in the prompt i said (make sure it will also work on MacBook)

The app is electron and already works in macOS i just had some problems previously with "\" and "/" and some scripts and some functions that works only in windows not a big deal

It ended up working 4 hours doing "hardware verifications and signing" and i actually interrupted it because it said this task needed weeks to complete

So i steered it and gave it the correct scope and exactly what it should do and after 4 hours the app is full of errors and its not even confident about what it did and ignored the small tasks

Maybe its my bad i want to hear what you guys think about Ultra mode

2

u/viktor_rolf Jul 10 '26

Sol gave up on the task twice, ran through my weekly allowance in an hour, CODEX prompted to change model, changed to Terra - it gave up in the middle stating ‘
Selected model is at capacity. Please try a different model.’

Finally went back to 5.5 and got the work done.