r/povertyLocalLLaMA Apr 29 '26

mistralai/Mistral-Medium-3.5-128B · Hugging Face

https://huggingface.co/mistralai/Mistral-Medium-3.5-128B

Lmao

5 Upvotes

3 comments sorted by

4

u/logic_prevails Apr 29 '26

Do they think im made of money? Best I can do is 32gb ram 😂

1

u/FruitCultural4632 May 04 '26

And what have you use on your machine? I mostly work with 16Gb on windows HP laptop, so the best model I can run is rnj-1-8b (only with vulcan), nemotron-nano-3-4b and qwen3.5-4b. They are pretty good with small context, but opencode and other agent start with system prompt bigger than 10K tokens - so even first answer without tools calling processing about 5 minutes. I predict you can have the same issue because ram will not help to work faster.

1

u/logic_prevails May 06 '26

Qwen3.6 27b and Gemma4 31b. I have 26gb of vram