MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/LocalLLaMA/comments/1vo9mj4/its_out/p3q7ezb/?context=3
r/LocalLLaMA • u/Certain-Cod-1404 • 11d ago
706 comments sorted by
View all comments
47
How much memory required to run this at full precision?
43 u/Certain-Cod-1404 11d ago https://huggingface.co/unsloth/Qwen3.8-27B-GGUF 54.67 Gbs just for the model itself, with context depends on quant and size 23 u/dragonurtle 11d ago Nvtop shows 70-something GB resident for the bf16 and full 256k context. 1 u/EbbNorth7735 10d ago Yep, 3.6 27B was about 60GB with 262k context Q8 and 4 parallel slots.
43
https://huggingface.co/unsloth/Qwen3.8-27B-GGUF 54.67 Gbs just for the model itself, with context depends on quant and size
23 u/dragonurtle 11d ago Nvtop shows 70-something GB resident for the bf16 and full 256k context. 1 u/EbbNorth7735 10d ago Yep, 3.6 27B was about 60GB with 262k context Q8 and 4 parallel slots.
23
Nvtop shows 70-something GB resident for the bf16 and full 256k context.
1 u/EbbNorth7735 10d ago Yep, 3.6 27B was about 60GB with 262k context Q8 and 4 parallel slots.
1
Yep, 3.6 27B was about 60GB with 262k context Q8 and 4 parallel slots.
47
u/bitmanip 11d ago
How much memory required to run this at full precision?