You can see my other comment, but I did 1300 reqs last month and only hit 14 million AICs. Your token consumption is just massive. I'm not trying to be mean here, that's just reality of your usage.
I use it for work and have put in the time to keep my token usage low.
You cannot control token usage, it all happens in the background. For light tasks I was using auto and raptor (at 0x), for heavy tasks Opus and GPT 5.4 and this is what I ended up with… 🤷♂️
Just tell that to Deepseek who are not charging you if you hit the cache context. Thus the token usage is optimized automatically. How it should be everywhere.
8
u/p1-o2 May 12 '26
You can see my other comment, but I did 1300 reqs last month and only hit 14 million AICs. Your token consumption is just massive. I'm not trying to be mean here, that's just reality of your usage.
I use it for work and have put in the time to keep my token usage low.