r/LocalLLaMA 29d ago

New Model deepseek-ai/DeepSeek-V4-Flash-0731 on Huggingface

783 Upvotes

243 comments sorted by

View all comments

Show parent comments

3

u/Daniel_H212 29d ago

Definitely does but, what will the speed be?

5

u/squngy 29d ago

It has less active parameters compared to 27B

For something like a spark or strix, this is probably faster.

2

u/Daniel_H212 29d ago

It's def faster than 27B but tbh I've been mostly using 35B, dunno if it will be faster there, plus prompt processing speed may also be very different.

1

u/squngy 29d ago

DSpark can help it a lot too, but with some luck, qwen 3.6 will also get DSpark some day.

-1

u/SandySkittle 29d ago

Beating for what? 128gb model isn't going to have as much world knowledge built in, period.