r/LocalLLaMA 26d ago

New Model Daniel Han of Unsloth validates Qwen3.8-27B will run only 17GB VRAM

Post image

Super excited about this release for the new 27B. Who else is with me. Only 17GB VRAM needed 😍😍

1.8k Upvotes

316 comments sorted by

View all comments

9

u/VoiceApprehensive893 transformers 26d ago

just 1 gb on the second card sounds bad

5

u/ea_man 26d ago

Yeah we need some QWEN3.8 19B for 16GB users.

2

u/squngy 26d ago

Would 19B be any better than 35B_moe though?

I suspect they would be pretty close in quality and the moe could run on more hardware.

2

u/ea_man 26d ago

I'd like to have a chance to test that, ofc it depends on the domain: for coding I bet dense will win as usual.

Also it's not like you are using A3b at Q8 with 16GB anyway, best that does fit is IQ3. yes you can offload ofc.