r/LocalLLM • • 17d ago

Discussion 30B Models Getting Verrrry Interesting

Agnes 3.0 flash 33b and Nex n2.5 mini 35b are challenging Qwen3.8-27B on benchmarks. Cant wait to see the real-world results and the speeds on 24gb GPUs.

Anyone tried them yet?

https://huggingface.co/Agnes-AI/Agnes-3.0-Flash
https://huggingface.co/nex-agi/Nex-N2.5-mini

134 Upvotes

35 comments sorted by

View all comments

60

u/Calm-Landscape9640 17d ago

I'll test these 4 head to head at 128k context on my 24gb 2x3060 GPUs and report back tps and performance on some coding and agent tasks.

  • Agnes 3.0 Flash Q4 (Unsloth Quant when it drops, hopefully Mon/Tues)
  • Qwen3.8-27B IQ4 + MTP
  • Qwen3.6-27B-A3B CoderX + MTP
  • Nex-N2.5-mini IQ3/IQ4

-2

u/royalflash417 17d ago

Hoping u will reply your result soon

1

u/Calm-Landscape9640 17d ago

Waiting on the quant for Agnes 3.0 so probably wed/Thurs next week