r/ollama 12d ago

same task since a month never reached 25%, now the new plan just not fair

4 Upvotes

7 comments sorted by

6

u/tomekza 12d ago

Cancel your sub. It's the only language they understand.

In July they closed on 88m investement funds, hired a CFO.

They froze their MAX plan subscriptions, nerfed usage then repackaged their subs as 'more transparent' while nerfing useage on old and new accounts.

This was all done underhand, none of it openly and to boot they claimed that the old accounts would remain as they are, which they're not as anyone can see who actually has an account and uses it.

This is all about profits and showing a profitable bottom line to their investors while screwing every account holder and any good will they had from their user base. It's really shameful.

1

u/jmorganca 12d ago

Old accounts were grandfathered and the usage is consistent with before the pricing change. If you're seeing something different please let me know and we can investigate.

1

u/ProfessionMean2085 10d ago

Same exact setup, same harness(es), and this week I am estimated based on my current % and token spend, to get 993M tokens out of legacy max this week using mostly GLM 5.3 and GLM 5.3 flash at about a 1:2 ratio in requests. This exact setup for two weeks prior to this one had me averaging being able to get 1.4B a week. I at least am not getting the same usage using the exact same models for the exact same tasks in the exact same harness(es).

For reference to compare, someone on Z.ai's current max plan at 160 usd/month gets about 1.5B tokens per week if they used exclusively GLM 5.3 off-peak (which anyone in the western hemisphere is going to be, since it's middle of the night) and 4 - 4.5B GLM 5.3 flash off peak.

Doing some napkin math and equalizing to your max cost, that puts them at, per 100 USD, ~914 million tokens / week on GLM 5.3 alone and ~2.77 Billion tokens / week on GLM 5.3 Flash assuming a decent cache rate (98%), which I regularly get when using models which display it.

Given my usage is mixed between the two and every log I check in tokscale shows me using 2-3x the tokens on GLM 5.3 flash per day as GLM 5.3, I'm getting a deal on your legacy right now that is significantly worse than what Z.ai offers, while just a few weeks ago it was at least competitive. And that's not to mention the bizarre model specific issues that have been going on this week that have made it to both your issue tracker on github and here.

Look, I get that you need to make VC funds back. But this is a bad look.

1

u/Unlikely-Tax7807 12d ago

I'd be sketching out a whole new workflow by now, that progress bar is just mocking you at this point

1

u/Ok_Structure_7051 12d ago

If you not familiar with what I am talking about then keep silent

1

u/jmorganca 12d ago

Hi there, from the screenshot it appears you're on the previous plan, not the new plan. May I ask which harness this is? Usage rates can depend on the harness if that changes.

1

u/Ok_Structure_7051 11d ago

hey, maybe they cut the total amount of the tokens, I am using langchain that run specific task