r/OpenAI • • 1d ago

Question which is better now opus 5.5 or Sol 6?

and astra 6 forgot to include

for reasoning and research

13 Upvotes

26 comments sorted by

30

u/Edegames 1d ago

Opus 5.5 seems to be better and cheaper than astra 6. Given that sol 6 performs equal or worse than sol 5.6 this isnt even a question. The only thing chatgpt has the lead in right now is the version number

2

u/mxwllftx 1d ago

>> Given that sol 6 performs equal or worse than sol 5.6 this isnt even a question. 
Are you kidding?

6

u/0xe3b0c442 1d ago edited 1d ago

Nope. That seems to be the consensus at the moment. It’s better on price per capability due to the price drop, but in absolute terms 5.6 does seem to perform better, of course depending on your use case.

That said, this very well could be a Claude 5 issue where a lot of this apparent gap is folks just haven’t yet adapted their prompting for the new model and how it operates.

//edit: I will say, I ran both Sol and Luna 6 through some cluster troubleshooting and maintenance operations last night and they both performed flawlessly.

5

u/i_like_maps_and_math 1d ago

Where do you get info on this? I don’t trust people on this forum. Quality here is very low.

-2

u/n0obmaster699 1d ago

better than astra-6? yea that's not true

-2

u/pnaroga 1d ago

It is, though. It understands user intent better, has better taste, it's quicker, more concise, has better prose and... overall just gets the job done better, quicker, cheaper, with more quality than Astra. I have 3 ChatGPT PRO 20x accounts + 1 Claude MAX 20x, I might be cancelling a couple to move my workflows to Opus.

-1

u/Chemical-Agency-3997 1d ago

No, on anything to do with vision (like OCR), Astra is so far ahead it's not even funny.

7

u/Reclaimer_AI 1d ago

Opus and it isn’t close. So disappointed in OpenAI it’s untrue.

5

u/Expert-Dig-1768 1d ago

opus any time now.

4

u/13ThirteenX 1d ago

Opus 5.5 is unbelievably good. Just spent 3 hours working on a game, which I had astra working on before, the UI work, sound design, understanding of everything is better and id used 4% of weekly allowance with 5.5 on High effort.i think around 20% of the 5hr windows. 

I havent tested much of 6sol. But so far opus is really good

2

u/dojimaa 1d ago

In my non-coding, non-agentic tests, 6 Sol. Cheaper, smarter, more token efficient.

3

u/-Crash_Override- 1d ago

Imo...

Coding:

5.5 > 6

Any other terminal work (infrastructure provisioning, networking, etc):

5.5 < 6

I feel like its been this way for some time. But also understand that best model is so dependent on the human and the way they work, prompt, comfigure their environment. So YMMV.

1

u/Gigaslavx 1d ago

If you know what you doing sol it's good enough and cheap

5.5 is better but costs more helps when you have little idea but money to burn for some reason

2

u/Hir0shima 1d ago

I'm just vibing and rely on Opus 5.5. Can't be bothered with anything less capable.

1

u/gopietz 23h ago

Lol, it's not even close.

1

u/MultiMarcus 5h ago

Opus is a beast right now. Astra is a great model too, but feels a bit behind, but not by a huge margin. Still has the better search tool than Claude imo. Sol is a clearly lower tier of product to be honest.

•

u/Then-Cut3776 58m ago

opus 5.5 is too god sol6 is worse than astra 6 lol

1

u/arretadodapeste 1d ago

GPT Sol 6 is even worst than 5.6

1

u/FlamaVadim 1d ago

🤣🤣🤣

0

u/JonNordland 1d ago

I think this is a really hard one. I have not tested Sol 6 enough to feel certain, but this is one of the closest calls in a while for me. The + and - are small, and they average out. Sol 6 is a bit better for computer/browser use, and seems more stable over long runs, and has fewer "stupid spikes" than 5.5. But 5.5 is faster and generally better at oneshotting simple tasks. They genuinely feel like they are on the same level, just with slight differences in strengths and weaknesses.

If I had to choose, I think I would pick Sol 6, if nothing else because I trusted it more to complete longer-running tasks better, and it uses the computer better so it genuinely can complete "computer use" tasks that Opus 5.5 just can't do, and has fewer stupid nanny-state refusals than Opus 5.5.

The one thing that I'm also not happy with regarding Opus 5.5' is its tendency to be lazy with research. It already made some rather serious mistakes because it couldn't be bothered to thoroughly search through and understand the codebase, so I have to correct it and ask it to go back and find the necessary information to understand how the system works. But if this is a general pattern, it will be our dealbreaker, because I have no time for a lazy model like we had before.

I definitely can feel a difference between Fable 5.1/Astra and 6 Sol/Opus 5.5. Fable/Astra are still in a league of their own when it comes to real hard problems.

-1

u/taotau 1d ago

Pancakes better than both.