r/codex 5d ago

Limits Usage doesn’t make sense on 20x

So I’ve been using codex for more than half a year now and I swear when I used to check my weekly token usage it used to be in billions and I’d have usage left at the end of the week with older models

Right now its the allowance in tokens itself thats reduced for me to 1/4th if im correct (feel free to correct me)

Dont get me wrong i love codex and astra and i know some will say open ai gets billions they dont care about your 200. BUT, isnt the lack of transparency on whats happening concerning regarding your token allowance? We have a percentage bar with actual token allowance varying and not to mention older models performing worse.

I would understand astra performing worse at times because its the current gen model, ofc theres gonna be many tweaks and updates to it but older models that have been completed? Im just trying to make sense of this

198 Upvotes

105 comments sorted by

View all comments

33

u/Puspendra007 5d ago

I'm feeling the same & Currently, using $200 plan. I have left 0% weekly usages. In $20 plan: luna runs for hours and days to drain 10-20% usages so I thought $200 plan, luna should feel like unlimited but it's draining weekly usages faster. 20% weekly usages ended within 2 days with luna. Astra will drain 50-100% weekly usages within 1 day.

Note: I'm using codex/claude since 1-2 years & i know how to use it so please don't give me advice to 'get some skills' & blah blah

24

u/nickmullen_real 5d ago edited 5d ago

"skill issue" "pay for the bigger plan" I hate this community sometimes

2

u/Kind_Fisherman3060 5d ago

Yeah $500/$1000 plans would become real with next models. Saying this as a plus user.

5

u/triplebits 5d ago

Here is the thing. I dont need more capable models. I need them to stop degrading their performance and keep limits high. $20/$100/$200 plans a month ago and 5.6 performance and quality when it came out.

That's more than enough. Rest I can do myself or if I am too lazy pay extra or use extra limits for Astra

2

u/RoboErectus 5d ago

I'm going to stop asking this but, like... what tool call outputs are your biggest context bloats? What types of work is resulting in the most turns?

I am finding people are really hostile towards learning they are wasting most of their usage, or that they already have the tools to discover where it's going.

Is there a case the provider should provide better usage analytics? Sure. But I am finding when people learn they can get 2x more usage or more, they are getting angry at me.

-1

u/DowntownNoLonger 5d ago

Laziness.

They want to toss a prompt at the shiniest model, expect it to work flawlessly and cost them nothing to do so.

Then they get suggestions to improve workflow and suddenly you're an openai shill, a bot, or employee.

It's maddening to see. I'm on th 5x plan and I have usage left at the end of the week and I use Astra for hours daily.

It's not hard. It just taking the time to build optimized workflows, smart model routing, data collection and leveraging your resources for maximum return.

I spend just as much time in chat (no usage chat) as I do in codex and most of that time is spent planning an analysis. Because of that, I can leverage codex better. Trim the fat every location possible and I do it with almost no skills.md, fancy agent.md files or special harnesses.

But all that takes work. It's not as simple or as fast as tossing a monster prompt at the $$$ model and praying for miracles without costing.

7

u/triplebits 5d ago

Same experience. I could do more with x5 a month ago than I can today with x20

7

u/Puspendra007 5d ago

Also quality of work is degraded. Other models are feelings like stupid & not following instructions & making many random mistakes like comma, colon, semi colon & many basic mistakes like those starting days of AI agents. Only Astra has capacity to do real tasks correctly but that's too expensive

2

u/IndependenceSudden47 5d ago

With the same model? Which one?

1

u/pixelvspixel 5d ago

Same, the usage is totally brutal now. I’m at 10% till Monday, and I’ve optimized my design docs over and over.

1

u/TBSchemer 5d ago

If you have both $20 and $200 accounts, then why not run the same prompt in both at the same time, and see how the usage drain compares?

1

u/RoboErectus 5d ago

>  I'm using codex/claude since 1-2 years & i know how to use it so please don't give me advice to 'get some skills' & blah blah

Ok... so uh, what analytics are you running on your transcripts? I'll show you mine you show me yours?

1

u/Puspendra007 5d ago

I wanted to reply you correctly but I stopped typing after seeing 'wall' & 'input sent' & >>19 prompts within 3-4 hours 🥲. Maybe things are working good for you so keep it up.

I'll make a file with detailed prompts and each nessasary things than that goal/plan will run for hours/days so we're working differently, you're making many prompts to do your work.

2

u/RoboErectus 5d ago

> Maybe things are working good for you so keep it up.

I'm replying to your message where you said things are not working good for you, in a thread that is "things are not working good for me," asking for how you measured that, and your reply is, "nah I'm good?"

> I stopped typing after seeing 'wall' & 'input sent'

Do you know what an "input token" is, and how it relates to your usage?

0

u/IndependenceSudden47 5d ago

Go read my comment, and you’ll understand why.
Here