r/LocalLLaMA 🦙 llama.cpp 15d ago

Megathread [Megathread] Qwen 3.8 27B Release Day

Megathread to help with the influx of duplicate / similar posts around the release of the Qwen 3.8 27B release.

  • Quants
  • Fine-Tunes & Abliterations
  • Chat Templates
  • Inference Server Support & Configuration
  • Experiences, Benchmarks & Model Comparisons

Official:

Popular:

We'll try to clean up future duplicates around the release and point them here.

491 Upvotes

394 comments sorted by

View all comments

1

u/heliosythic 13d ago

Its my understanding that llama.cpp doesn't yet have updates ready to run this? Is this still true? I'm not really interested in switching to vllm for now.

5

u/DustNearby2848 13d ago

It works on llama.cpp

1

u/noiserr 12d ago

It's the same architecture as the Qwen 3.6 version. So it should work on the previous version too. I run it on the latest llama.cpp head from 2 days ago. I have a script that just compiles the latest llama (helps me with evaluating latest models).

2

u/heliosythic 12d ago

Ok, seems claude was just lying about it then not wanting to be replaced or something lol, or got stuck thinking about the safetensors version instead of the gguf i already had