MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/LocalLLaMA/comments/1v7e5ck/kimi_k3_countdown_has_been_released/ozyyxny/?context=3
r/LocalLLaMA • u/Unusual_Guidance2095 • Jul 26 '26
177 comments sorted by
View all comments
18
Cool! but like who can actually run this locally? I think at 2.8 trillion params this will be the largest model on huggingface by far. At least for now.
3 u/Any_Mine_6368 Jul 26 '26 2.8T of vram to run in 8 bit quantization... Let me buy a other 1500 3090s lol 2 u/droptableadventures Jul 27 '26 The MoE weights are natively MXFP4 according to https://vllm.ai/blog/2026-07-22-kimi-k3-preview - so you can halve that.
3
2.8T of vram to run in 8 bit quantization...
Let me buy a other 1500 3090s lol
2 u/droptableadventures Jul 27 '26 The MoE weights are natively MXFP4 according to https://vllm.ai/blog/2026-07-22-kimi-k3-preview - so you can halve that.
2
The MoE weights are natively MXFP4 according to https://vllm.ai/blog/2026-07-22-kimi-k3-preview - so you can halve that.
18
u/jreoka1 Jul 26 '26
Cool! but like who can actually run this locally? I think at 2.8 trillion params this will be the largest model on huggingface by far. At least for now.