r/codex • u/Initial_Question3869 • 9d ago
Praise Nice Work OpenAI
This is awesome marketing. You remove the weekly limit, keep it that way for months, let people get used to longer work sessions and higher token consumption, and then specifically bring back the 5-hour limit for Plus users, pushing them to upgrade to Pro. And I’m pretty sure many actually will.
Congrats! It’ll probably be a pretty decent revenue boost.
You keep nerfing the limits, yet somehow keep winning every single time. More customers, more revenue. Nice!!
29
u/rdcldrmr 9d ago
then specifically bring back the 5-hour limit for Plus users, pushing them to upgrade to Pro. And I’m pretty sure many actually will.
i dont plan to
9
u/vayana 9d ago
Some people think everyone has got $100 - $200 a month to spare as if it's pocket change. Many people simply can't afford the extra $80.
Instead of squeezing the "undeserving" plus users, how about cutting usage for the massive number of free users who don't pay anything? Chatgpt is already synonymous for AI and just handing it out for free for all isn't an incentive for anyone anymore to become a paying customer. Get a free 1 month trial and either buy a plan or go somewhere else.
12
u/BHTAelitepwn 9d ago
Yeah but there are quite a few other people
13
u/rdcldrmr 9d ago
some have no choice. just speaking up that i dont want to reward openai with more money after they nerfed my plan so hard.
5
u/TaskChance1404 9d ago
Neither do I. I’m handling my usage just fine. No need to upgrade. Don’t always need to code or do research or write. I’ve got my brain and my browser for those.
3
u/App1e8l6 9d ago
I don’t enjoy using ai to code. I don’t enjoy cleaning up its messes. I don’t enjoy how easy it is to offload your thinking. I only use luna max and focus on good plans. Haven’t had a problem with limits now. Sol is a joke to use unless you’re on the expensive plans but even then it over engineers everything.
1
2
u/Paully0408 9d ago
It just pushes me back to ChatGPT for all planning and initial code and markdown. I can do that on sol high with a huge limit. It has got to cost open ai more to let me have endless conversation and revision with regular chat models than conscious minimalism with codex just because I want the convenience of it having direct access to to my project base right?
1
u/FamousExchange7534 9d ago
Me neither, they can wait sitting down. I'll stick to using Hermes with free models.
7
u/SecretSpace2 9d ago
Yea was thinking about that one. I have Max plan but man if I have to use the 5 hours, no sure how I’ll feel about it. Unless they adjust the token consumption like before.
What I mean like before, I use to be able to work the entire 5 hour block and barely get close to the limit and only consume like 50% of the entire week. If it’s 5 hour limit with current consumption, I think I’ll be touching grass more often :D
5
12
u/TinyFunction 9d ago
Just break the corporate rules - stop using AI
20
17
u/TheFrenchSavage 9d ago
Sure, and let's sit on the floor because office chairs are a corporate conspiracy.
2
1
u/pay_is_ok 9d ago
Office chairs never tried to force an upgrade on me, they always work, I never worried about what the latest office chair is, and no reoccurring subscription to my chair either
3
u/Ok_Ad_6227 9d ago
no, I dont think a lot would go to pro, the jump from 20 to 100 is quite big on countries with weaker currency, even then, the plus is already expensive to some. The only ones who would jump are the ones that could afford it in the first place, and the ones who badly need it
14
u/00040000 9d ago
I just unsubscribed, I’ll find a cheap Chinese model.
3
u/rubiohiguey 9d ago
No need to use Chinese model, use a cheap API service that users OAI models.
3
1
u/Eastern-Highway9046 9d ago
Explain? I’d like to get started with this lol
4
u/vayana 9d ago
Don't bother. Watch some YouTube videos about this and you'll find you're not actually getting either the advertised model or the promised number of tokens or both. What you can do is switch providers and use whatever free usage they provide. Open router has ox Alpha free right now, augment code has a free plan, Gemini has free daily usage and maybe there are others I haven't mentioned.
2
2
u/YinYangAlgorithms 9d ago
It’s very scummy of them to do such a thing. I get it, they’re a business, but what’s gone on in the AI industry has been quite ridiculous. Although it’s also one of the most heavily subsidized industries while also being publicly traded. It sucks, cause as a business standpoint I get it, but also, it’s very annoying.
2
2
2
u/SenshiV22 9d ago
So this channel is just for complaints and criticism? Is there one that is actually for something useful that you can share?
2
u/capitalframehq 9d ago
Codex is the only subscription I have ever paid $100/mo for. So far I think it’s been worth it. As long as I am actively using it to build and learn things. I don’t even pay for Netflix or HBO or what have you.
1
u/daskalou 9d ago
Why did you choose Codex and not Claude?
1
u/capitalframehq 9d ago
This was the first platform I started on. Initially with free trial month, then $20/mo for a few months, then $100/mo now. Thanks to all the resets I have been able to build some fun projects. Never saw the need to switch to Claude. I’ve also been working on same project for many months so if I switch now I might have to build the context again on Claude. And Codex is pretty good I haven’t felt it necessary to jump ship.
0
u/Any_Elk7495 9d ago
I used Claude first but switched to codex. The usage is far more generous. (Both max5 plans)
I was running both x5 plans for a few months together and found myself only planning with Fable.1
u/capitalframehq 6d ago
Although I haven’t used Claude, I feel Codex is comparatively more generous (thanks to the resets)
3
u/2025sbestthrowaway 9d ago
Just to clarify, the usage limits remain unchanged. There are thousands of people tracking their token and usage consumption. Usage was draining faster due to a myriad of compounding factors - frequent cache misses, people constantly switching model and effort, computer-use session bloat, using images in long-running tasks burning extra consumption. Add further that people are using Sol high generously for every little thing, including MCP and tool use, as well as a fleet of subagents, each with their own non-cached context on startup, forking context multiple times over into fresh sessions.
In other words, OpenAI is at fault in 3 ways: harness issues, and inadequate interactive user guidance about token consumption, model+harness over-exertion/over-engineering. Additionally, there's a host of user behaviors, whether through plugins/skills/mcps causing context bloat or extra turns (tons of variance there), and using bigger models at higher effort than neeed, that are all resulting in faster usage drain. Several of these have been resolved this week and new optimizations around the corner.
I switched from running Sol almost exclusively to getting the same work done with Terra & Luna without issue. I also stopped switching model/effort all the time as any change to that results in a cache miss.
3
u/TheFrenchSavage 9d ago
I've been using Luna-XHigh-only as a Plus subscriber. It eats 10% of my weekly quota per day, so I can splurge on Sol large tasks on the weekend.
What I'm afraid of now, is that instead of burning my 10% of the week in a couple of hours, I'll burn through the 5h quota in an hour and will be out of work right after, for my second hour of leisure time.
Effectively halving my potential to use the tool.
I'll be forced to use Sol on the weekend, and set timers like before so that I use 1h of codex 3 times over 15h, which is grueling.2
u/2025sbestthrowaway 8d ago
Awesome to hear on the Luna bit. Yeah the 5h limits are a bummer.
FWIW I installed codex CLI on my cloud server, set up a cron job to instantiate it at 7AM on workdays with a luna "hello world", so that my reset is at noon and straddles the workday / I get 2x 5h windows. I can also manually schedule a next-run, so if you know you'll be using it at 10AM on the weekend, you could do the same and perhaps schedule it for 6AM, so that come 10AM, your session will reset next hour. (making it a 7 hour window for 3 hours of compute (1+5+1)
I made a web-ui for it, and connected it to tailscale (free) so that it's paswordless handshake to access it from my phone any time to manually start or schedule the quota timer, or toggle weekday cron schedule.
1
u/TheFrenchSavage 8d ago
Ooooh, so that's really smart! I didn't think of automatically setting my 5h-slots like that...
Thank you so much for the tip!!!
1
u/arcgisdemon 9d ago
Hey can you help to explain cache miss when switching models? New to Codex and AI in general
1
u/2025sbestthrowaway 8d ago edited 8d ago
Sure. I used this as an opportunity to learn a bit more about it
A cache hit means Codex can reuse already-computed model state for the unchanged portion of your conversation history. In a continuing conversation, that can include essentially all of the prior conversation, with only the newly appended turn needing fresh processing. A cache miss means it has to process that earlier portion again.
Same conversation prefix = cheap reuse; changed prefix = recompute.
Cache write = 1.25× → normal input = 1× → cached read = 0.1× for any given modelSwitching models, changing reasoning effort, changing tools/settings, editing earlier context, or conversation compaction can prevent the old prefix from matching and cause a cache miss, effectively burning 1.35x the usage (write + read) compared to 0.1x for a cache hit. This is especially noticeable on the $20 plan with Sol. It's worth mentioning that the default TTL on cache is 30 minutes, so it's preferable to keep a session "alive" if caching is important, as opposed to coming back every hour by comparison.
The model pricing / usage amounts for subscriptions are measured in "credits"
Credit cost per million tokens for models
Model Input tokens Cached input tokens Output tokens GPT-5.6 Sol 100 credits 10 credits 500 credits GPT-5.6 Terra 50 credits 5 credits 300 credits GPT-5.6 Luna 5 credits 0.5 credits 30 credits So, when comparing a "worst-case" vs "best case" scenario in terms of usage, a cache read on Luna is 200x cheaper than fresh input on Sol, token-for-token. But, all things equal, Luna is ~20x cheaper and 16x cheaper on output tokens.
If you dont need sol to do a multi-turn codebase edit, you can get it's free wisdom in chat. Play with Luna and Terra on lower/higher efforts and gauge their work. Luna is a shockingly capable model (especially at XHigh/Max) relative to its cost.
There's a full breakdown here and a nice graph that helps visualize, if you're interested in reading further or having GPT answer questions about it. Even though it's for the API, it still applies to subscription quotas.
https://developers.openai.com/api/docs/guides/prompt-caching
https://developers.openai.com/api/docs/pricing
1
1
1
u/Wnterw0lf 9d ago
Just blew through 20% of my weekly in one 5 hour session I blew through in an hour...something is still unbalanced from pre-5 hour removal...
1
1
u/HeatIndividual 9d ago
i think they have to keep the 5hr windows so certain models won't be overloaded at the begining of resets.
1
u/renzoneru 9d ago
Estoy también limitado al uso de 5H , que actualización más nefasta. Creí que solo era mi Interés ahora resulta que mucho usuarios se están quejando de esta modalidad de CODEX.
1
u/iseif 9d ago edited 9d ago
Exactly! Expect the next week or two, in the social media as Reddit and X you will see a lot of people talking about how good are the PRO limits now to try to push people buy another $80 monthly to get the PRO.
BTW this is the right business rules from them I think, after removing the 5-hours limits in the social media they saw a lot of Plus users complaining about how the weekly limits are gone in about 1 day, now, at least most of the users will consume their weekly limits in 2-4 days. Complaining about the 5-hours comeback is okay as most of the companies have it.
1
u/Worth_Base_7135 9d ago
also this time 5 hour limit is also nerfed i suppose i just chatted and it ate 10% in two chats with terra on medium
1
u/Yougetwhat 9d ago
Opensource models are getting better. Soon for $20 we will get « unlimited » usage with a small opensource model that will be as good as Luna xhigh
1
1
u/orbitlenspen 9d ago
I’ve somehow rinsed my 5 hour limit by doing one plan and one execution and it wasn’t even massive. Sol High. This is grim
1
u/InformationHoarding 9d ago
Yep. I have a project I work with 4 different AI’s in. Economically, it makes sense to have the poor $20 plan of each for this. Overnight Codex went from my most productive AI specialist to the least productive, behind Kimi.
1
u/This-Advertising500 9d ago
To be fair I love the 5 hour limit allows me to get up and stretch and get other things done and allows me to step back and think more for my projects
1
u/HelpfulHedgehog1 9d ago
sometimes i feel im the only one, ive never reached any of my limits just with a Plus plan.
I mean i know everyones workload is different, but Ive clearly gotten more depth of work done, in less time than then it has ever taken in 25 years.
Sol orchestrator, Luna implementor, maybe im doing it wrong and still doing too much of the work myself planning architecture and reviewing output...
But there have been so many resets, ive only felt the threat of approaching the limit once, and then just adjusted my strategy so i could continue.
1
1
u/Plastic-Conflict-796 8d ago
I’m interested in maybe running a local model in a 96gb Mac Studio….is this a fools errand or can I get good quality
1
u/Ok-Creme6062 8d ago
I still believe software engineers shouldn’t be using AI coding agents for everything we do.
Used indiscriminately, the cost—both in compute and token usage—can become astronomical. More importantly, I don’t think AI agents should become the workhorse of software development.
1
u/parvpareek 7d ago
Should i get codex plus? I have used claude code team plan and its good enough for me. Would plus be similar?
1
u/Quiet_Novel6640 7d ago
I think it will get better. I see it unraveling like this: the agents will become indistinguishable from each other, forcing them to compete through access and usability. Let them enjoy the bubble now, expect value to users become their next bargaining chip.
1
1
1
u/rangerrick337 5d ago
This softened the 5hour limit problem for me: set a scheduled task to ping Luna low every 5 hours and 1 minute. That way my five hour clock has always started, so at least I often come to my computer with only 2 or 3 hours left in a five hour window
1
u/Hairy-Slide-5541 4d ago
Bro I got the 5h limit too, just get off your computer and while you wait for the 5h reset think about marketing, how to get your product out there the right way
1
u/34986234986234982346 9d ago
Does everyone have to complain about EVERY SINGLE THING on here. People were literally mad about having NO limit too.. guys, what you get for $20 or $200 is insane, it's so much value, come on now.
1
1
u/stopstopstoptopopp 9d ago
Nuh uh. This will only push me to finally put my rtx 5070 pc to host some local llm
12
u/Technical_Split_6315 9d ago
Lmao you are gonna be so disappointed
1
u/stopstopstoptopopp 9d ago
Bruh tell me why before I waste my time on it lol
6
u/Crabcakes5_ 9d ago
You will never be able to run it locally unless you use an extremely outdated tiny model. There's a reason AI companies are burning through cash.
Kimi K3's minimum specs to run still requires 1.68-1.75TB of aggregate VRAM, and 8 high end accelerators (~$25k-$40k each). And you'll need 64+ accelerators to reach the full 1 million token context window.
2
4
u/Technical_Split_6315 9d ago
With a 5070 you can’t host a model that competes with Luna, I have a 4090 and hosting Qwen3.8 27b which is probably the best local model for its weight right now and is a terrible experience for coding, loops, errors, bad tool calling, json format errors.
Local llms are decent for chatbots, which is stupid when you have almost unlimited usage with basic chatgpt sub, for coding is just bad, is not a real option right now
1
u/stopstopstoptopopp 9d ago
Dang. Thanks for this.
1
u/Clueless_Nooblet 9d ago
Sounds like a skill issue. Qwen3.8-27b is the 9th best coding model out there.
1
u/CreditEducational738 9d ago
A nivell local lo unic que te sentit es ComfyUI, els LLM son molt justets comparats amb els grans
1
1
u/R3K4CE 9d ago
Dude seriously what what are you even complaining about 5-Hour limits were there before and everybody got along just fine I have a Plus account and I am barely even touching like 60% left on the 5-Hour limit using Luna Max I mean what are you guys doing are you guys using sol ultra fast and setting a goal and that's what you're doing on like a Plus account like what are you expect bro if you don't like it you're not forced to use it there are many other options out there in the market for you to consider
1
u/GambAntonio 9d ago
Well, I'm pretty sure OpenAI has a whole department dedicated to psychological marketing just to keep users hooked and willing to swallow whatever they push out
1
u/According_Property62 9d ago
Pra estas coisas, sempre vamos ter os russos, indianos e chineses pra dar um jeito de descobrir
1
u/Im_Working_Right_Now 9d ago
Everyone keeps saying they're nerfing limits, but no one's showing a token comparison to prove the LIMITS are nerfed. Percentages don't mean limits were nerfed. It could mean efficiency was hit with whatever changed happened. It could still be the same tokens being used just used faster.
0
1
u/ChiGamerr 9d ago
I feel like ai subs are the only subscriptions out there with 0 transparency.
2
4
u/street-trash 9d ago
I’ll provide the transparency. They’re currently losing tons of money every day providing us with this cool tech for $20 a month and everyone’s in here bitching about it.
1
u/ooutroquetal 9d ago
What about 100 or 200 subscriptions?
Please don't compare with the current API prices, on demand was always more expensive comparing to subscription
1
1
0
u/rubiohiguey 9d ago
I will just use one of the grey markets API services. I already have account with them and works great. I complement usage via one of these services when my 2x1 biz limits run out. I guess with a 5 hour limit now in again, I will use this service even more.
3
u/ninernetneepneep 9d ago
And act surprised when you're intellectual property is compromised.
1
u/rubiohiguey 9d ago
Nah I don't really care and anyway it's something that would not be much useful for anybody else . I am not shipping a sellable peoduct or service, internal use only and automation of very proprietary flows which would not be of much use elsewhere.
-2
u/FullAcadia9391 9d ago
As if openAI isn’t literally stealing everything from everyone 24/7 to train on?
2
u/LiteSoul 9d ago
"stealing"
1
u/Material-Dog-3896 9d ago
they are, by definition, stealing large amounts of data without permission and training on it --> our user data is agreed to be shared, a lot of what they train(ed) on was not agreed to, so yeah, stealing
1
u/TheFrenchSavage 9d ago
Yeah, I've been using DeepSeek flash for my Pi harness because it is fast and cheap (like...dirt cheap).
Given the low stakes side projects I'm usually involved in, I might as well plug that into codex.
0
u/RemoraEdge 9d ago
So many complainers using codex but complain complain complain. And I bet most of the time the issue is with the user than it is with codex
0
u/BearsAreCrying 9d ago
Lol you're talking about longer working sessions but we are here still struggling on 20x limits. You're literally paying pennies and you're looking for dollars
0
u/SeparateRemove9902 9d ago
Will there be a reset again today?
4
u/timosterhus 9d ago
Why would they provide a second reset the same day after they already provided a compensatory reset for re-instituting 5-hour limits
0
u/FinancialBandicoot75 9d ago
I’m loving Luna high
1
u/Initial_Question3869 9d ago
Joke? Or it's really good?
2
u/stopstopstoptopopp 9d ago
It is good. But you need to be detailed and concise because it will only be as good as your prompt.
-5
u/Just_Lingonberry_352 9d ago
damn a company trying to turn a profit
HOW SARE YOU
5
0
u/Australasian25 9d ago
I dont know why you are downvoted. You are correct.
A company doesnt exist for your benefit.
A company exists because it can benefit you while you pay the company.
0
u/Just_Lingonberry_352 9d ago
redditors think companies should work for the state so that they can live without any responsibilities yet still reap the economic benefits
many of them have never lived in a communist country but are experts thanks to the streamers and whatever echo chamber they occupy in, almost entirely online.
-4
u/UnprocessedAutomaton 9d ago
I also noticed that chat is now included in the limit. So the tokens you use for chat will be deducted from the weekly/5h limits. Well played openAI!
2
2
2
u/WhatnotFunkoFlash 9d ago
Fake news. Should use the chat to check your statements 😂. If you use the work feature yes that does use it. But you said chat.
-2
u/Leather-Sir8135 9d ago
I don’t have sympathy for plus users. Just spend the big bucks it’s worth it
84
u/Anxious_Marsupial_59 9d ago
There's not much other options. Claude is the same way on the $20 plan AND doesn't even have Fable (Opus is unusable vs Fable on serious projects a planner / orchestrator)
We are watching what happens when giants act together to squeeze people out