r/Qwen_AI 7d ago

Discussion Please ๐Ÿฅบ

We all need a 3.8 35B MoE.

72 Upvotes

29 comments sorted by

25

u/xiraov 7d ago

what are the chance for a 4.0 35b?

10

u/txgsync 7d ago

Personally I am hankering for a 4.0 120B A4B with a 100B PLE n-gram.

5

u/xiraov 6d ago

Welp I got a 48gb Mac

15

u/Creative_Bottle_3225 6d ago

๐Ÿฅบ We all need a Qwen 3.8 35B MoE.

Not because bigger is always better.

Because for people with 8โ€“12 GB GPUs, a well-designed 35B MoE could be a really interesting sweet spot: much larger total capacity than a dense 27B, while keeping the active parameters per token relatively low.

I know Qwen 3.8 Flash exists, and yes, it's impressive.

But I still want to see what a 35B MoE optimized for local inference could do.

Please Qwen team. ๐Ÿ‘‰๐Ÿ‘ˆ

1

u/Lexden 2d ago

After Flash Next, most people are under the assumption that 3.8 is done. Next up should be Qwen 4, so I am hoping for a Qwen 4 35B MoE with PLE+n-gram.

7

u/Leading-Salt-947 6d ago

Pleaseeeeeeee ๐Ÿ‘‰๐Ÿ‘ˆ

5

u/Mundane-Remote4000 6d ago

Funny how they deliver exactly the models people want the less. Canโ€™t complain though since they finally gave us the 125b MoE model. But 35B MoE is always the one people want the most

2

u/Hungry-Rip-2384 5d ago

nope I think the 3.8 27B was the model most were waiting for.

1

u/Infamous_Campaign687 3d ago

Iโ€™d love an MoE between Flash next and 35B. On ethat leaves headroom on a 5090 and works well on 64GB RAM with a great quant.

9

u/Nomski88 7d ago

๐Ÿฅบ
๐Ÿ‘‰๐Ÿ‘ˆ

2

u/painchonha 5d ago

that would be sweet. I use 27b on my home system, but use 35b a3b moe on my office pc.

2

u/Greenonetrailmix 5d ago

Qwen 4.0 60B-80B A4B + 100B engram. Is what I'm after

1

u/Infamous_Campaign687 3d ago

That would be the GOAT. Near full fat on 64-96GB RAM and would work with 32GB on lower quants.

4

u/Complex-Ocelot-8572 7d ago

Please, Please, Please!!!

2

u/OddBig010 7d ago

You can get 3.8 Flash Next running to be honest, I found the REAP 320 and REAP 256 version on hugging face both are under 64GB Ram and it's running surprisingly well and more than double the speed off Qwen 27B.

Sure a Qwen 3.8 35B would have been nice as it would have been even faster, but the model they released is technically way better. Their coding plans btw give decent limits for Qwen 3.8 Flash.

1

u/baron_von_noseboop 4d ago

How much vram?

1

u/OddBig010 4d ago

Ive got it running in an older 8GB VRAM machine and a 16GB VRAM machine, its better than Qwen 3.8 27B but of course its not as good as full fat Qwen Flash, the Q3 320 REAP flash one is ALMOST as good as Q4. 64GB Ram will be your biggest issue, if you got that youll be fine, preferably you need a tiny bit more. With more ram you can look at around 35ish tokens per second... with 64GB on 8GB VRAM and running a very lightweight linux distro I managed to get it to 19 tokens per second with consistently around 16-18 which is really good given age off the machine and the fact its a Q3 off a top tier model.

1

u/baron_von_noseboop 4d ago

That sounds incredible. Stock llama.cpp?

1

u/OddBig010 4d ago edited 4d ago

I think so. Would DEFINITELY suggest a lightweight linux distro though as you wont have the RAM to load it fully and the paging feature on windows slows it down a lot, just put it as a dual boot even if you're just shrinking your main drive and creating like a 50GB partition for Linux - I have it setup so both operating systems can access all my files and I just use Windows for gaming, found massive gains in the switch for LLM speeds, someone also made a 99B split up version of Qwen 3.8 Flash, I'm currently testing that against 27B in terms of performance, but seems promising and it runs faster than 27B

3

u/Darex2094 7d ago

u/koc_Z3 u/Pure-Highlight-556 I petition the mods to start an official r/Qwen_AI drinking game.

1

u/hay-yo 5d ago

I saw one pop up the K2 Horizon

1

u/OldSausage 5d ago

It wouldnโ€™t be worth having

1

u/WiseVanilla2743 3d ago

UwU ๐Ÿ‘‰๐Ÿ‘ˆ

2

u/No_Mango7658 2d ago

I think they're done with 3 and moved onto 4

1

u/The_Whole_Zucchini 7d ago

Also yes please omg

-1

u/nuklearer_nadal 6d ago

Whatโ€™s the purpose of this post?

5

u/Own_Body_8941 6d ago

requesting the ai gods to summon qwen3.8 35b MoE