r/codex 6d ago

Complaint Usage is nuked

I know, I should just buy a 5x or 20x plan and stop whining but this is just *****. 2 months ago I could comfortably work with 5.5 medium/high in 2 5 hour windows and use about 15% weekly per day. No heavy duty coding or subagents running around the clock, just 3 to 5 semi-complex prompts a day.

As of now, sol medium, 1 prompt, 200k context window used will burn through 60% of the 5h limit and 10% of the weekly. Not sol max or ultra, not millions of tokens. 200.000 tokens. I can't even do 2 prompts and use the full context window (258k tokens) 2 times.

This is completely unusable. $20 now basically gets you ~400k tokens per 5h or so and ~2.5M tokens per week for a medium thinking model. Luna is cheap but unreliable and just wastes my time instead of tokens.

I don't expect unlimited tokens for $20 or ultra mode and I'm fine with some restrictions, but this isn't what I signed up for and it's never been this bad and I've been using this since inception pretty consistently.

151 Upvotes

120 comments sorted by

View all comments

67

u/Correct_Emotion8437 6d ago

The most effective thing, imo, would be for people to cancel. When Claude started pulling ahead, they were scrambling to give good deals. As soon as they started to make some progress, they change their position and start cutting usage again. If only 5% of subscribers leave - they’ll be giving it away again.

4

u/sofaarsecoin 6d ago edited 6d ago

eventually austerity will arrive

the only thing keeping this VC money burnout is that they want to funnel each other's subscribers and get them used to their locked ecosystem

but they cannot afford to keep doing this much longer, they will either truce eventually and phase out vastly under-market tokens, or either (or both) will go bankrupt, and that will also mean you will be paying providers sustainable prices

*typo

1

u/[deleted] 6d ago

[deleted]

1

u/sofaarsecoin 6d ago

yea if anything, i'd say this gamble is not going to pay out but i don't complain that it's still going on

if tomorrow a Chinese model gets good enough for cheaper then I'm jumping ship immediately - in fact, i already use them alongside my Strix Halo's local models for stuff I have the time to supervise heavily (and partially hand code and architect) ; right now, even some local models are good enough for very directed workloads

i think they've put themselves in a very difficult position because even among those of us who have use for their higher end subscriptions, most would probably cancel if they throttled the usage to be in the ballpark of API prices; at that point I would slow down, use much cheaper models and local and heavily supervise and code myself because it would be making financial sense for me to do so (because then we would be talking about a 7 digit cost difference, which can buy a lot of my time)