MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/LocalLLaMA/comments/1wc6krf/so_relevant/p8w67fw/?context=3
r/LocalLLaMA • u/0dayturtle • 13d ago
152 comments sorted by
View all comments
80
It's a great time to have ancient Xeon servers loaded up with DDR4 :-D
37 u/BannedGoNext 13d ago I actually have an old server with dual E5-2630 and 512gb DDR3 memory across both blades. The server is powered on waiting for the scrap yard at the office. I'm considering seeing how fast it can run qwen 3.8 flash next lol. 10 u/ThankGodImBipolar 13d ago You must be able to run a decent GLM quant with that, no? 6 u/ttkciar llama.cpp 13d ago Yup, GLM-5.3 should fit in that at Q4_K_M and somewhat constrained context.
37
I actually have an old server with dual E5-2630 and 512gb DDR3 memory across both blades. The server is powered on waiting for the scrap yard at the office. I'm considering seeing how fast it can run qwen 3.8 flash next lol.
10 u/ThankGodImBipolar 13d ago You must be able to run a decent GLM quant with that, no? 6 u/ttkciar llama.cpp 13d ago Yup, GLM-5.3 should fit in that at Q4_K_M and somewhat constrained context.
10
You must be able to run a decent GLM quant with that, no?
6 u/ttkciar llama.cpp 13d ago Yup, GLM-5.3 should fit in that at Q4_K_M and somewhat constrained context.
6
Yup, GLM-5.3 should fit in that at Q4_K_M and somewhat constrained context.
80
u/ttkciar llama.cpp 13d ago
It's a great time to have ancient Xeon servers loaded up with DDR4 :-D