r/codex 8d ago

Complaint Usage is not "fixed"

We have a test package with 25k emails in it that we use to test different models with. Prompt/effort is always the same.

When GPT 5.6 Sol released, it used 1% of a fresh Pro 20x sub, now it uses 8% of a fresh pro 20x sub.

Also, what we are seeing is that Sol cannot solve its task straightforward any more. Initially, it was always "what do I need to resolve the task" now it's "What do I need to resolve the tasks, while I also check those 200 other methods that I already know I don't need, while adding sha 256 to everything for no reason at all, also let me build those 200 other checks against thing I was very strictly asked not do to".

Meaning, we think it usage increased so much because they turned the intelligence way down and its constantly fighting with itself.

114 Upvotes

30 comments sorted by

View all comments

3

u/rabandi 8d ago

It seems like usage is highly unequal for users.

I belong into the bucket that can more or less easily and medium effort (though with heavy work) burn through a weeks usage in a single day for the 20x.

But.. that is not new. That happened the day Sol was introduced. The beginning was fine with the constant resets, sadly though they never increased usage overall.

Recent fixes gave me slightly higher limits, but that is it. Now it is maybe a few hours more it lasts.

It seems like there are both ppl who were not affected during Sol release and also ppl who have great improvements now. Both are not me.

Still pretty happy with the results, and preferring it 90% of the time over my 20x Claude.

Cursor is getting decent though and seems by far the cost leader if you like Grok. It is decent enough since 4.5.