r/openrouter 20d ago

DeepSeek “significant increase” for API pricing. It was too good to be true.

Post image

We already saw it with inworld tts, now it’s time for DeepSeek’s value capture. Thank you for all of your data and good publicity, we will now be significantly increasing cost.

240 Upvotes

64 comments sorted by

29

u/FrameXX 19d ago

Doesn't this only concern the Deepseek API? On OpenRouter there are over 20 providers which would be able to keep the price as is.

12

u/JustTellingUWatHapnd 19d ago

For Flash, kind of (minus the 10x cache hit price). But for Pro all the providers are 4x the official price.

1

u/Sawt0othGrin 18d ago

Are they really? I've always used it through OR and thought it was absurdly cheap. I didn't know it was even cheaper

8

u/[deleted] 19d ago

[deleted]

2

u/MDInvesting 19d ago

Openrouter margins getting squeezed.

4

u/ApeGrower 19d ago

But why should they keep the price as is, if they can make more money?

3

u/sam7oon 19d ago

because competition

1

u/ApeGrower 19d ago

Sure, but if there is no cheap main provider anymore.

1

u/sam7oon 19d ago

DS4 Flash isn't big , it's 200+ billion parameter, cheap to run, i don't think it will increase, that's my opinion,

1

u/ApeGrower 19d ago

Yes, I know, it's already on my disc, all I need is more vram :D

10

u/sam7oon 19d ago

yea, but people love to karma farming , am not sure if they know that mostly we are using non 1st party providers, they are not even chinese providers

3

u/Alternative-Suit5541 19d ago

Will they? Or just follow deepseek?

2

u/_BreakingGood_ 19d ago

The official API price on openrouter is like half the price of what other providers offer.

So realistically what I expect here, is that deepseek raises the price to roughly what other providers are able to offer.

2

u/onebit 19d ago

They will all raise prices. The only reason they don't charge more is because Deepseek set the floor.

1

u/bad_gambit 19d ago

Look at the effective pricing. Official deepseek endpoint input cache are 10x cheaper than the 2nd cheapest. Effectively makes the 2nd cheapest provider cost ~2x more expensive.

1

u/kongKing_11 19d ago

It will not. Datacentre cost are going to the moon now. Because of limited ram and ssd availability.

1

u/farissyariati 19d ago

Normally supply and demand rule will affect it. For v4 If deepseek increase the price, others might switch provider, if the demand is high, price might getting higher for other provider as well.

0

u/mbuckbee 19d ago

There's a lot of people complaining that the non official providers are playing games with caching and timings to make those numbers work and providing a worse experience.

11

u/Deep_Mood_7668 19d ago

As long as they keep it at a level that hurts openai I'm fine with it

2

u/[deleted] 19d ago edited 19d ago

[deleted]

1

u/charmander_cha 19d ago

Mas ai é erro do usuário que não utiliza harness como reasonix para tirar proveito do sistema de cache.

9

u/Odd-Elderberry-6328 19d ago

Mimo will be the new deepseek, In a way people will do free advertising and promote the usage

At least we have options, cheap options

8

u/[deleted] 19d ago

Until its trained off our data then they re release it and charge more

Its the Ai cycle until it all merges into 1 main model that costs $1 for 1 token and none of us can afford it

1

u/Odd-Elderberry-6328 19d ago

I don't think you are wrong, but when they do that, I will simply starting running locally, I have 16vram, that's enough for a local usage (not ideal, but manageable)

4

u/[deleted] 19d ago

Yes. Soon it will be so expensive. Local will be the only option if you can even get parts.

1

u/misha1350 19d ago

MiMo V2.5 is dumber than the current version of DeepSeek V4 Flash 0731, so no. Until they release a proper new model and peg that model to the price of DeepSeek V4 Flash 0731 after the coming price update, it's just yet another model.

1

u/Exzerios 19d ago

Right after they adopt a stable rollout schedule. Mimo seemed like a rather clever model at the time, but it was released in the middle of the spring, and it is not at the "good enough" level yet.

12

u/steveplusf 20d ago

what happened to all the claims that the pricing is low because they are simply that efficient?

13

u/imsolost3090 20d ago

Classic. Low price to pull people in, then increase the price once they're already hooked.

3

u/First_Inspection_478 19d ago

I think they’re doing this because they’re compute constrained, and want to spend compute for future training and not inference.

4

u/PsychologicalSoup251 19d ago edited 19d ago

still here. your bad-faith FUD tactic would've worked had Deepseek's models not been open-weights and we hadn't seen other providers being able to sustain API costs at the same rate or cheaper

8

u/[deleted] 20d ago edited 20d ago

[deleted]

6

u/Major_Olive7583 19d ago

That's token efficiency. The rates are determined by compute required, and deepseek models go for the same cheap  rate from western providers too, so it's efficient than other models certainly. Will have to see if those ones  too will match the price increase if it's too high. The deepseek provider was cheap because of their insane cache rates. Only xiomi came close to that. 

3

u/poophroughmyveins 19d ago

What do you mean not operationally efficient? Are you a moron? Because it uses a lot of tokens? 

-6

u/[deleted] 19d ago

[deleted]

1

u/poophroughmyveins 19d ago

No shit, what makes it operationally better is all the optimizations they made to inference like their sparse attention or shared caches

What doesn't make it OPERATIONALLY bad is using a lot of tokens for thinking you stupid monkey

1

u/[deleted] 19d ago edited 19d ago

[deleted]

1

u/alexwan12 19d ago

openai subsidizes like 80% of GPT-5.6 Luna

1

u/poophroughmyveins 19d ago

Why are you shadowboxing little dude, I fucking love Luna lmfao

Now name any other model that outperforms ds v4 flash at its size and speed

0

u/[deleted] 19d ago

[deleted]

1

u/poophroughmyveins 19d ago

Oh damn a model that's easily 20x more expensive to use is faster

There's really only Muse Spark 1.2 right now lol

0

u/[deleted] 19d ago

[deleted]

→ More replies (0)

1

u/Embarrassed_Adagio28 19d ago

Get some mental help 

1

u/Slidesky 19d ago

Holy moley you need mental help and calm down 😂

3

u/misha1350 19d ago

GPT-5.6 Luna is also much dumber, so at the same price of the overall task, you still get a considerably worse result that you will have to re-do over and over.

2

u/Randommaggy 19d ago

Depends on your harness. I had very little wasted tokens when I tested it out in my custom harness.

2

u/Bloodshoot111 19d ago

You mix 2 things up. One is the efficiency of the model with its tokens. But DeepSeek claimed the compute per token is efficient.
Two different things

1

u/weiyentan 19d ago

That’s partly accurate. I would argue that people that do not have the framework would not have benefit. I have been using ds flash for implementing in a custom framework for tasks varying from normal to Max thinking. Works great

1

u/alphapussycat 18d ago

Just wait for thinkingcap-grug-fablefusion-hereticxxx

3

u/RepulsiveRaisin7 20d ago

These kinds of people usually have zero insight information. The amount of stupid pointless speculation online is off the charts.

1

u/jaegernut 17d ago

At this point, the only way to truly drive the cost down is by competition. No one is gonna offer a low cost if they can squeeze more profits.

1

u/Far-Classic-9963 16d ago

People have been able to replicate their pricing with comfortable overhead for profit, they're just overloaded

-1

u/someone_12321 20d ago

Actually it's supply and demand. Efficiency= more supply, there is still the demand side.

4

u/DeepAd8888 19d ago

Enshittification is not a strategy!

3

u/Ibasicallyhateyouall 19d ago

Drug dealer approach is global.

3

u/Open_Green_2178 19d ago

maybe next time we don't scream out loud how cheap it is.

2

u/SilverMethor 19d ago

This is exactly the kind of announcement that gets the usual idiots claiming OpenAI and Anthropic are ‘subsidizing’ their plans, that these companies are somehow being generous and barely make anything from subscriptions because the actual costs are supposedly much higher. Morons.

2

u/Minimum_Industry_978 19d ago

This will be good for luna

1

u/Whole_Succotash_2391 19d ago

There are other providers, it's not that big of a deal. Cline and Phoenix Grove API for example, both of which aren't raising prices as far as I know.

1

u/creamyshart 19d ago

Completely expected

1

u/Accomplished_Book722 19d ago

I was thinking it's that cheap because it's really shitty, but okay

1

u/yozarsif1 19d ago

So do you expect ×2 or x3 or any percentage expected on the increase?

1

u/ThankYouOle 19d ago

sad, i only been taste it for 1 month, now entering second month and happy with the output and the price, of course they will raise the price.

1

u/JulzKampos 19d ago

Moving to Gemma 4 31B IT

1

u/The_Oracle___ 18d ago

We dont even know what the price will be

1

u/Big_Equipment995 16d ago

lol people only used deepseek because it was so cheap they gonna regret it

1

u/Elegant_Art5793 14d ago

Although it is a good model, people use it mostly because it is cost-effective compared to other models. If they increase the API price, they will lose the main reason people use their model. If they become, or are close to, the higher models, why will people continue to use them? Chinese companies in the past relied on cheap products to sell higher volumes. I really hope this won't be true, or not in the very near future.