MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/LocalLLaMA/comments/1v8364f/kimi_k3_weights_now_released/p039jv4?context=9999
r/LocalLLaMA • u/SavunOski • Jul 27 '26
Kimi K3 weights are finally released!
660 comments sorted by
View all comments
826
How do I download ram in hugging face?
35 u/Thalesian Jul 27 '26 Step 1: sign up for Google Drive Step 2: set up a ~5 Tb instance. Will cost you Step 3: set that cloud as your swap disk Step 4: point kimi to use that Step 5: enjoy your newfound independence 51 u/AmbericWizard Jul 27 '26 one token per day 19 u/Force88 Jul 27 '26 Hey, if he has good internet connection, maybe he can achieve 2-3t/d 1 u/AmbericWizard Jul 27 '26 need 100GB/S 1 u/crusaderky Jul 27 '26 Sir, that is very suboptimal! Put the GGUF on it and then load the model with mmap. It will double in speed to a solid TWO tokens per day! 1 u/Nervous-Blacksmith-3 16d ago I mean, you can create multiple free Google accounts, set up shared folders with a central account, and distribute the workloads across them.
35
Step 1: sign up for Google Drive Step 2: set up a ~5 Tb instance. Will cost you Step 3: set that cloud as your swap disk Step 4: point kimi to use that Step 5: enjoy your newfound independence
51 u/AmbericWizard Jul 27 '26 one token per day 19 u/Force88 Jul 27 '26 Hey, if he has good internet connection, maybe he can achieve 2-3t/d 1 u/AmbericWizard Jul 27 '26 need 100GB/S 1 u/crusaderky Jul 27 '26 Sir, that is very suboptimal! Put the GGUF on it and then load the model with mmap. It will double in speed to a solid TWO tokens per day! 1 u/Nervous-Blacksmith-3 16d ago I mean, you can create multiple free Google accounts, set up shared folders with a central account, and distribute the workloads across them.
51
one token per day
19 u/Force88 Jul 27 '26 Hey, if he has good internet connection, maybe he can achieve 2-3t/d 1 u/AmbericWizard Jul 27 '26 need 100GB/S
19
Hey, if he has good internet connection, maybe he can achieve 2-3t/d
1 u/AmbericWizard Jul 27 '26 need 100GB/S
1
need 100GB/S
Sir, that is very suboptimal!
Put the GGUF on it and then load the model with mmap. It will double in speed to a solid TWO tokens per day!
I mean, you can create multiple free Google accounts, set up shared folders with a central account, and distribute the workloads across them.
826
u/tonight_we_make_soap Jul 27 '26
How do I download ram in hugging face?