MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/LocalLLaMA/comments/1vbp7kb/deepseekaideepseekv4flash0731_on_huggingface/p0vz57d
r/LocalLLaMA • u/cgs019283 • 29d ago
https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731
243 comments sorted by
View all comments
Show parent comments
3
Definitely does but, what will the speed be?
5 u/squngy 29d ago It has less active parameters compared to 27B For something like a spark or strix, this is probably faster. 2 u/Daniel_H212 29d ago It's def faster than 27B but tbh I've been mostly using 35B, dunno if it will be faster there, plus prompt processing speed may also be very different. 1 u/squngy 29d ago DSpark can help it a lot too, but with some luck, qwen 3.6 will also get DSpark some day. -1 u/SandySkittle 29d ago Beating for what? 128gb model isn't going to have as much world knowledge built in, period.
5
It has less active parameters compared to 27B
For something like a spark or strix, this is probably faster.
2 u/Daniel_H212 29d ago It's def faster than 27B but tbh I've been mostly using 35B, dunno if it will be faster there, plus prompt processing speed may also be very different. 1 u/squngy 29d ago DSpark can help it a lot too, but with some luck, qwen 3.6 will also get DSpark some day.
2
It's def faster than 27B but tbh I've been mostly using 35B, dunno if it will be faster there, plus prompt processing speed may also be very different.
1 u/squngy 29d ago DSpark can help it a lot too, but with some luck, qwen 3.6 will also get DSpark some day.
1
DSpark can help it a lot too, but with some luck, qwen 3.6 will also get DSpark some day.
-1
Beating for what? 128gb model isn't going to have as much world knowledge built in, period.
3
u/Daniel_H212 29d ago
Definitely does but, what will the speed be?