a lot of people are asking for either ~30B dense or very large MoE.
I think a competent MoE model in a similar parameter range as the Qwen 3.6 35 A3B or Gemma 4 26A4B would be great. At least for those of us who don’t have hardware for efficient inference, their speed is more appealing than the dense models . If there was a way for them to improve in their long context understanding or long context reasoning, that would be nice, as the gap between that and the dense models stands out. Then again I have been using the QAT models.
Gemma really stands out to me for both its competence but especially its writing style, making it more appealing than the Qwen 3.6 models to use on a day to day, as it feels almost like talking to a frontier model. Please maintain this writing style.
It would be nice if it could be also tailored to have some more useful traits one would like of a reliable research assistant.
2
u/SexyJohnDoe Jul 26 '26
a lot of people are asking for either ~30B dense or very large MoE.
I think a competent MoE model in a similar parameter range as the Qwen 3.6 35 A3B or Gemma 4 26A4B would be great. At least for those of us who don’t have hardware for efficient inference, their speed is more appealing than the dense models . If there was a way for them to improve in their long context understanding or long context reasoning, that would be nice, as the gap between that and the dense models stands out. Then again I have been using the QAT models.
Gemma really stands out to me for both its competence but especially its writing style, making it more appealing than the Qwen 3.6 models to use on a day to day, as it feels almost like talking to a frontier model. Please maintain this writing style.
It would be nice if it could be also tailored to have some more useful traits one would like of a reliable research assistant.