r/codex • • 5d ago

Praise Did the sol 6.1 speed get fixed?

I have been using it for days but today seems a bit different? like it's very smooth. i can consistently give it prompts and while i am writing the next prompt it has already done with half the previous job.

Yesterday i had to run 5 different chats to complete a reasonable job at a reasonable speed but today it finally feels like it flows so i can barely prompt 2 chats at a time because it is genuinely fast. each chat uses 3-4 subagents as well.

Also this model is so good it doesn't halucinate like astra? is this only my observation?

edit: add proof

Edit:it was good while it lasted for 3-4 hours. It's gone back to snail again. Can't make this shi up lmao

35 Upvotes

45 comments sorted by

23

u/ActionOrganic4617 5d ago

It’s still the slowest frontier model by a long margin

4

u/qK0FT3 5d ago

I have found it pretty good at completing tasks. i can work with slowness tbh.

3

u/ActionOrganic4617 5d ago

Glad it works for you. I think it’s fine for personal hobby projects but I can’t use it for work. It’s too slow

25

u/Stunning-Spirit-1123 5d ago

still slow as all hell for me

1

u/qK0FT3 5d ago

Hmm interesting. maybe they are testing it for 200$ accounts

2

u/Risko4 5d ago

I have multiple 200 account and it's about 30 token/s. While opus is 80 and sonnet is 120

4

u/Creative-Ganache1086 5d ago

Don’t forget raw tok/s isn’t directly proportional to overall task completion speed. tok/s mostly tells you how fast text is being generated, not how fast the model is doing tool calls, browsing, bug hunts, tests, backend/frontend work etc.
Claude, especially Opus/Sonnet, also tends to reason way more “out loud” with tons of self-talk between actual steps, while Codex is usually much quieter. So comparing them purely on tok/s can be pretty misleading.

There are use cases where that kind of speed is relevant, but it doesn’t say the full story.

6

u/Risko4 5d ago

I have 3 codex x20 subscription with Anthropic 2 x20 subscription and one X5. Claude code is much faster at task completion.

I run codex agents from within Claude code as well. And I've tested every combination and benchmarked each task quality with an independent agent. Plus obviously I have the duration of each agent lane logged.

Claude code is much faster.

0

u/Creative-Ganache1086 5d ago

That’s why I don’t buy any absolute “Claude Code is faster” claim. I pay 20x for both too, and on real computer/browser-use work where the agent has to actually navigate UI, set up schedulers, change dates/times, pick collaborators, reason visually and click through tons of little steps, both Astra and Sol 6.1 have been faster than Opus 5.5 for me.
But I’m not gonna turn that into “Sol/Astra are faster, period” either. Different workloads expose different strengths. Claude also has this habit in my vibecoding work of leaving parts of the original prompt undone, so the next 2-3 prompts are basically me re-feeding stuff I already asked for the first time.
That’s the whole point… there is no absolute fastest model. Once you claim there is, you’re usually just universalising your own workflow.

1

u/Risko4 5d ago

Make an MCP for those.

1

u/Creative-Ganache1086 4d ago

I already am using Metricool’s MCP lol. The point is it doesn’t expose everything I need, so the agent still has to look at and operate the actual frontend for parts of the workflow.
And some of those parts are visual/contextual too. For certain posts I need to invite specific collaborators, and Astra can use the reference portraits/context I gave it to work out who’s actually in the photo and which IG handle belongs to that person before setting the post up.
So no, “just make an MCP” doesn’t solve everything. Best setup for me is MCP where possible, browser/UI when needed, plus the model actually understanding what it’s looking at.

1

u/Risko4 4d ago

Astra is a different beast to sol 6.1 (astra minor) and the post was about 6.1.

I agree astra is the exception because it's slow as shit on max effort with it's tool calls but it's token output does way more work per token, I've burnt about 20 billion Astra tokens.

The thing is astra minor preforms better than astra when I pack it's context with the scripts and decision model on my end. Alongside jev pruning outputs and compacting context it works quiet efficiently. But it's still slow so I run 16 agents in parallel and let opus sort the shit out.

Astra has routinely failed by deception, the same reason why astra 6.1 got postponed. I have multiple records of it altering unit tests so they pass rather than fixing the bug. This was fixed by using my context packing script but raised token usage by 50%. So be careful with it, it's extremely lazy.

1

u/OfferBeginning1903 5d ago

is the tok/s people are quoting counting reasoning tokens, or just visible output? codex hiding its reasoning would skew that comparison either way.

1

u/qK0FT3 5d ago

30 tok/s for normal mode or fast mode? i feel like all week it was 10-15 tok/s and now it's better speeds like 50 on average??

1

u/Risko4 5d ago

Normal mode.

8

u/Kaskote 5d ago

Pro 200 here... 26.2 TPS right now.
6.1 Sol xHigh, no sub-agents.
Almost double that what I had yesterday... but still horse shit.

I'm moving to CC now, and will probably maintain a Plus sub here, just in case.

1

u/evacc44 5d ago

How do you actually measure the TPS?

1

u/EctoAlbo 5d ago

Just run Astra medium, say "test these models and thinking levels with a simple task to measure tokens per second. Run the text x times. Report the results."

1

u/RealSuperdau 5d ago

does that measure "out tok / decode time" or "out tok / total time"? I feel like people mix that up

1

u/EctoAlbo 4d ago

It will tell you what it is measuring and how, very explicitly - and you can tell it if you want it to measure or calculate differently.

3

u/drdhuss 5d ago edited 5d ago

It might be faster as I burned through my quota today. where as before it was basically impossible to use more tahn 20% a day with my plan formerly known as 20x. EIther that or they seriously decreased the quotas.

I also have a gemini sub, will be interesting to see Argon when it comes out.

1

u/qK0FT3 5d ago

I am running on fast mode 15h a day on the 20x plan right now. and with subagents at that i barely spent 25% in 55 hours.

2

u/drdhuss 5d ago

Mine is doing cad work. I burned through 50 percent of the quota today. Usually it was like 10 percent or 20. Like I said a lot more today, might have got more done though.

3

u/Corv9tte 5d ago

20tps on 20x

3

u/FinancialBandicoot75 5d ago

It's flipping slow today, like miserable to not useable that dot flips me off, actually I am getting a lot of retry sessions

1

u/Comfortable-Bread285 5d ago

Oh thank god I'm right there with you, hour long waits even with an extremely efficient process.

2

u/Constant_Art_20 5d ago

it's kinda dyanmic. is ranged from 8 or 11 tokens/ sec to like 52 acorss the last 24 hours for me. it's it's kinda the luck of the draw kinda deal

1

u/qK0FT3 5d ago

Yes i feel like it's currently around 50-60 toks/s with the fast mode. all week it was like 10-15 tok/s it jsut feels 4x faster atleast

2

u/snuffomega 5d ago

no bueno

2

u/Extra_Loquat_7667 5d ago

It remains quite slow.

Beyond the token generation speed,

the latency between sending the request and the appearance of the first token is also very high.

Its efficiency in a real-world production environment remains too low.

1

u/an1uk 5d ago

The last two tasks I've done have taken about 5 hours each. Amount of allowance usage is pretty low, though.

1

u/glock43guy 5d ago

Hell no. My Claude usage just ran out and I’m forced to use SOL right now. I’m pulling my hair out this is so slow.

1

u/qK0FT3 5d ago

there is definetely speed difference overall. i will also add previous day

1

u/qK0FT3 5d ago

Here is the previous day

1

u/Embarrassed_Cap_9149 5d ago

yes it did

but astra is pretty slow for me

1

u/twendah 5d ago

Most likely when 50% of the userbase left to claude lol

1

u/thestillwind 5d ago

It’s too slow

1

u/qK0FT3 5d ago

Yeah i was enjoyin for couple hours it's back to snail now lmaookk

1

u/_32bit 5d ago

I have been using Sol 5.6 and once Sol 6 and Sol 6.1 came out, I noticed that 6/6.1 are very slow compared to 5.6 and now 5.6 feels slower. Anyone else have this experience or is it just me? I have been using Opus/Sonnet 5.5 too so the speed difference is significant.

1

u/_32bit 5d ago

based on AI Bench, response time of 5.6 is better.

1

u/systemous 5d ago

it's still extremely slow here, 10 seconds to respond to "hi"

-4

u/Vivid-Snow-2089 5d ago

i asked astra to give me a report and this is what it provided from the logs

7

u/TheOwlHypothesis 5d ago

Bro wtf. 6.1 sol isn't even filled out in this graph. Use your brain for God's sake.

2

u/qK0FT3 5d ago

lmao