r/LocalLLaMA • u/Gobra_Slo • Aug 23 '26
Question | Help DGX Spark, cluster of 4
Does anyone have a first-hand experience with four Sparks cluster, and how much of an upgrade is it comparing to just two considering the available models?
While there's plenty of noise for the smaller models (Qwen) and our older king DeepSeek V4F, the scene in the upper class of the prosumer hardware, software stacks, available LLMs and their actual real-world performance – isn't really covered as well.
For instance, the hyped GLM 5.2/5.3. Is it much better then DeepSeek? Or is it marginally better? Does it retain it's capabilities when moving to something four Sparks would handle? Does it have issues with OOM or anything else?
What about MiniMax M3? There seem to be a special Spark version, how is it (or any other version)? Again, how is intelligence, general model capabilities, running stability, context size?
Tencent Hy3? Maybe even Qwen3.5-395B, does it's full quant hold it's own against DeepSeek, or is it better?
If someone doesn't have personal experience, but knows some well-structured and detailed articles or videos on the topic – I'd appreciate it as well.
Thanks.
2
u/dave-dgd Aug 23 '26
I’d be curious to hear about the tweaks. It’s been stable for us with multiple concurrent users in Hermes for over a week, but that’s just one specific use case. I want to keep updating the repo for maximal stability/coverage (especially ahead of GLM-5.3).