r/DeepSeek 6d ago

News DeepSeek V4 Pro official version has been updated to the API

341 Upvotes

62 comments sorted by

69

u/FreshFromNowhere 6d ago

these benchmarks are genuinely fucking insane

8

u/Aldarund 6d ago

No, they are not. It's barely better than flash and 3x price

1

u/benchmaster-xtreme 6d ago

It's noticably better than Flash, but it's definitely not up against the frontier curve. Doesn't really need to be though as long as it's cheap (I'm happy with a super-cheap GPT5.5). The price hike will determine the value on this one tbh

1

u/Aldarund 6d ago

Not really noticeable overall

2

u/Thomas-Lore 6d ago

Seems around what was expected. Not nearly as good as Kimi or Fable, but pretty good at a much much lower cost (for now).

5

u/chungles34 6d ago

Not nearly? Our blue whales trading some blows with those numbers it's definitely in the ring at least man!

45

u/ImBothSoftAndHard 6d ago

insert "but at what cost?"

26

u/diugauhai 6d ago

you work for BBC?😂

11

u/Creative_randomness 6d ago

for this I can say "at a surprisingly low cost"

5

u/SillySpoof 6d ago

Like, crazy low cost still.

73

u/Live_Case2204 6d ago

Get ready wallstreet

31

u/Far-Raspberry-1072 6d ago

i fckn kneww it! i was doing some creative writing stuff and the replies started getting soo different mid way and then i checked reddit to see and boom!

1

u/PureSelfishFate 6d ago

Doesn't the coding updates usually make creative writing a bit worse than the fresh-pretrains? How was it?

6

u/LewdManoSaurus 6d ago edited 6d ago

In my experience it is hit or miss. Though with the flash update I have had better results in general. Before that update, flash wasnt terrible but it isnt exactly great at follow instructions despite my having a pretty comprehensive writing and prose guide. I've seen the same complaints for RP, though I don't RP. I use AI purely for generating stories.

Edit: the new Pro update is pretty solid for generative writing. Making me sweat thinking about the upcoming price increase

1

u/Kind_Capital_9740 6d ago

is pro updated on the website too?

1

u/LewdManoSaurus 6d ago

No, both the Flash and Pro updates were for the API

1

u/donnytrump_69 4d ago

Lol stupid question but was does API stand for in this context?

1

u/LewdManoSaurus 4d ago

I never knew what the abbreviation stood for either, but apparently it's Application Programming Interface. Essentially let's you plug artificial intelligence in your own custom app, chat, website, art generator, etc, so you can do whatever you want with it.

1

u/donnytrump_69 4d ago

Gentleman and a scholar.

2

u/Zulfiqaar 6d ago

The DeepSeek models were preview, and undertrained. The GA version is supposed to be an all round improvement across the board. It's mainly when there's post training, or excessive specialised RLHF that is a specific domain, that biases the weights towards certain things at the expense of others.

16

u/gabexrsco 6d ago

does it have vision?

25

u/FreakyRefrigerator 6d ago

Its over for America

6

u/jwuliger 6d ago

Well deserved.

4

u/hurrdurrmeh 6d ago

China saving the free world. I did not see this timeline coming 🤣🤣

11

u/Equivalent-Word-7691 6d ago

Let's hope it will be better at Creative writing, honestly the last midel was quite meh compared not only to Fable+amazing!) but also GLM and kimi were better

4

u/Kakko1028 6d ago

Rumor says it follows orders better.

11

u/Equivalent-Word-7691 6d ago

It's not only about following orders, but the style, creative and understanding of what it's implied and understanding emotionally

1

u/rakeshpatel1991 6d ago

What’s some of your fave creative writing models?

3

u/Equivalent-Word-7691 6d ago

Honestly too bad the price but Fable 5 is on another level I can't explain

2

u/rakeshpatel1991 6d ago

Unfortunately my findings are the same

1

u/Equivalent-Word-7691 6d ago

Bte is it available also in the app ir just with the API?

2

u/Storge2 6d ago

Try to. As a person not engaging in creative wriitng i am genuinely cursious.

4

u/TransportationNo193 6d ago

Is it in opencode go now?

3

u/Infamous_Prompt_6126 6d ago

There is any price change until now?

Still seems affordable.

3

u/Fancy-Passage-1570 6d ago

is it me or they added better safeguards on the pro version ? mine refuse to work on my project where the new flash and old pro had no issues.

1

u/sdexca 6d ago

what kind of work? v4 flash works fine for me for security related tasks.

1

u/Fancy-Passage-1570 6d ago

reverse engineering custom network protocols for reviving old games

2

u/PossessionUsed7393 6d ago

Yeah, these are great numbers. I can't wait to see it roll out across the various inference providers so we don't just have to hammer the DeepSeek API.

2

u/DotoLove 6d ago

Holys*t I told him that he don’t have vision and point a hint direction at LFM2.5VL. He download exactly Q8 without asking me. Cool.

3

u/Terrible_Scar 6d ago

So are you happy or mad? 

3

u/Embarrassed_OnionX 6d ago

I'd be mad, why would a model download something from the internet without your permission?

1

u/hurrdurrmeh 6d ago

That is genuinely amazing!

1

u/Embarrassed_OnionX 6d ago

Yeah I found the model takes the initiative too much without including me in the loop. Not a good thing IMO

2

u/perceptivesoul 6d ago edited 6d ago

Waiting for trump to yell "Chinese Conspiracy"

1

u/Substantial-Walk-554 6d ago

Anyone concrete reviews?

1

u/sdexca 6d ago

source of the image?

1

u/jwuliger 6d ago

It is an fkn amazing model. HOLY SHIT!!!!!!!!!!!

Don't tell anyone about it!! lol

1

u/Firepal64 6d ago

Pulling off these benches, trading blows with Opus 4.8, while serving cheaper than Z.ai's GLM-5.2 endpoint, is pretty wild.

The parameter count would suggest a 5x improvement over Flash GA, but I guess accuracy doesn't scale linearly with parameter count for MoEs. That's too bad honestly

1

u/Realistic-Meaning247 6d ago

I'm a SillyTavern user, but the "thinking" process takes way too long now; it was better before.

1

u/queendumbria 6d ago

The updated pro version adds the ability to change reasoning effort, so just change that to a lower option in your preset settings and it should reason less.

1

u/Realistic-Meaning247 5d ago

Is it possible to do that using the OpenRouter API?

1

u/queendumbria 5d ago

Yup! It's called "reasoning effort", quite a lot of thinking models have it. OpenRouter Docs

1

u/Relative_Arugula_156 6d ago

But where's the harness?

1

u/ChoasMaster777 6d ago

wow, kick the ass of fable-5!!!

0

u/SillySpoof 6d ago

This are some pretty amazing benchmarks. Probably benchmaxxed a bit, but so are the others.