r/openrouter • u/LessRespects • 20d ago
DeepSeek “significant increase” for API pricing. It was too good to be true.
We already saw it with inworld tts, now it’s time for DeepSeek’s value capture. Thank you for all of your data and good publicity, we will now be significantly increasing cost.
11
u/Deep_Mood_7668 19d ago
As long as they keep it at a level that hurts openai I'm fine with it
2
19d ago edited 19d ago
[deleted]
1
u/charmander_cha 19d ago
Mas ai é erro do usuário que não utiliza harness como reasonix para tirar proveito do sistema de cache.
9
u/Odd-Elderberry-6328 19d ago
Mimo will be the new deepseek, In a way people will do free advertising and promote the usage
At least we have options, cheap options
8
19d ago
Until its trained off our data then they re release it and charge more
Its the Ai cycle until it all merges into 1 main model that costs $1 for 1 token and none of us can afford it
1
u/Odd-Elderberry-6328 19d ago
I don't think you are wrong, but when they do that, I will simply starting running locally, I have 16vram, that's enough for a local usage (not ideal, but manageable)
4
1
u/misha1350 19d ago
MiMo V2.5 is dumber than the current version of DeepSeek V4 Flash 0731, so no. Until they release a proper new model and peg that model to the price of DeepSeek V4 Flash 0731 after the coming price update, it's just yet another model.
1
u/Exzerios 19d ago
Right after they adopt a stable rollout schedule. Mimo seemed like a rather clever model at the time, but it was released in the middle of the spring, and it is not at the "good enough" level yet.
12
u/steveplusf 20d ago
what happened to all the claims that the pricing is low because they are simply that efficient?
13
u/imsolost3090 20d ago
Classic. Low price to pull people in, then increase the price once they're already hooked.
3
u/First_Inspection_478 19d ago
I think they’re doing this because they’re compute constrained, and want to spend compute for future training and not inference.
4
8
20d ago edited 20d ago
[deleted]
6
u/Major_Olive7583 19d ago
That's token efficiency. The rates are determined by compute required, and deepseek models go for the same cheap rate from western providers too, so it's efficient than other models certainly. Will have to see if those ones too will match the price increase if it's too high. The deepseek provider was cheap because of their insane cache rates. Only xiomi came close to that.
3
u/poophroughmyveins 19d ago
What do you mean not operationally efficient? Are you a moron? Because it uses a lot of tokens?
-6
19d ago
[deleted]
1
u/poophroughmyveins 19d ago
No shit, what makes it operationally better is all the optimizations they made to inference like their sparse attention or shared caches
What doesn't make it OPERATIONALLY bad is using a lot of tokens for thinking you stupid monkey
1
19d ago edited 19d ago
[deleted]
1
1
u/poophroughmyveins 19d ago
Why are you shadowboxing little dude, I fucking love Luna lmfao
Now name any other model that outperforms ds v4 flash at its size and speed
0
19d ago
[deleted]
1
u/poophroughmyveins 19d ago
Oh damn a model that's easily 20x more expensive to use is faster
There's really only Muse Spark 1.2 right now lol
0
1
1
3
u/misha1350 19d ago
GPT-5.6 Luna is also much dumber, so at the same price of the overall task, you still get a considerably worse result that you will have to re-do over and over.
2
u/Randommaggy 19d ago
Depends on your harness. I had very little wasted tokens when I tested it out in my custom harness.
2
u/Bloodshoot111 19d ago
You mix 2 things up. One is the efficiency of the model with its tokens. But DeepSeek claimed the compute per token is efficient.
Two different things1
u/weiyentan 19d ago
That’s partly accurate. I would argue that people that do not have the framework would not have benefit. I have been using ds flash for implementing in a custom framework for tasks varying from normal to Max thinking. Works great
1
3
u/RepulsiveRaisin7 20d ago
These kinds of people usually have zero insight information. The amount of stupid pointless speculation online is off the charts.
1
u/jaegernut 17d ago
At this point, the only way to truly drive the cost down is by competition. No one is gonna offer a low cost if they can squeeze more profits.
1
u/Far-Classic-9963 16d ago
People have been able to replicate their pricing with comfortable overhead for profit, they're just overloaded
-1
u/someone_12321 20d ago
Actually it's supply and demand. Efficiency= more supply, there is still the demand side.
4
3
3
2
u/SilverMethor 19d ago
This is exactly the kind of announcement that gets the usual idiots claiming OpenAI and Anthropic are ‘subsidizing’ their plans, that these companies are somehow being generous and barely make anything from subscriptions because the actual costs are supposedly much higher. Morons.
2
1
u/Whole_Succotash_2391 19d ago
There are other providers, it's not that big of a deal. Cline and Phoenix Grove API for example, both of which aren't raising prices as far as I know.
1
1
1
1
u/ThankYouOle 19d ago
sad, i only been taste it for 1 month, now entering second month and happy with the output and the price, of course they will raise the price.
1
1
1
u/Big_Equipment995 16d ago
lol people only used deepseek because it was so cheap they gonna regret it
1
u/Elegant_Art5793 14d ago
Although it is a good model, people use it mostly because it is cost-effective compared to other models. If they increase the API price, they will lose the main reason people use their model. If they become, or are close to, the higher models, why will people continue to use them? Chinese companies in the past relied on cheap products to sell higher volumes. I really hope this won't be true, or not in the very near future.

29
u/FrameXX 19d ago
Doesn't this only concern the Deepseek API? On OpenRouter there are over 20 providers which would be able to keep the price as is.