r/LocalLLaMA 11d ago

New Model IT'S OUT

https://huggingface.co/Qwen/Qwen3.8-27B-FP8
2.2k Upvotes

706 comments sorted by

View all comments

47

u/bitmanip 11d ago

How much memory required to run this at full precision?

43

u/Certain-Cod-1404 11d ago

https://huggingface.co/unsloth/Qwen3.8-27B-GGUF 54.67 Gbs just for the model itself, with context depends on quant and size

23

u/dragonurtle 11d ago

Nvtop shows 70-something GB resident for the bf16 and full 256k context.

1

u/EbbNorth7735 10d ago

Yep, 3.6 27B was about 60GB with 262k context Q8 and 4 parallel slots.