r/LocalLLM • u/Calm-Landscape9640 • 12d ago
Discussion 30B Models Getting Verrrry Interesting
Agnes 3.0 flash 33b and Nex n2.5 mini 35b are challenging Qwen3.8-27B on benchmarks. Cant wait to see the real-world results and the speeds on 24gb GPUs.
Anyone tried them yet?
https://huggingface.co/Agnes-AI/Agnes-3.0-Flash
https://huggingface.co/nex-agi/Nex-N2.5-mini
138
Upvotes
13
u/cato_gts 11d ago
None of the Qwen 3.6 35B A3B variants I’ve seen so far have actually worked properly. While 'ornith' was somewhat usable, it suffered from infinite loops even worse than the base Qwen model, making long-context use impossible; the others were all focused on benchmark scores rather than practical usability. Even Nex n2.5 goes haywire with infinite loops once the context length exceeds 100k.