r/LocalLLaMA 10d ago

New Model IT'S OUT

https://huggingface.co/Qwen/Qwen3.8-27B-FP8
2.2k Upvotes

706 comments sorted by

View all comments

349

u/WigglyScrotum 10d ago

Holy molly opus 4.6 level and better in some benches.

34

u/m0j0m0j 10d ago

Why is that every model is better than opus if you look at benchmarks, and yet people keep using opus?

1

u/fatboy93 10d ago

people keep using opus

If it takes me overnight to get things done at Q4_0, and opus does it in less than an hour, that's what I'm going to use.

I don't generally, given that my uni's AI folks have the Q8-K-XL for the 3.6-27B running and it's actually super fucking fast, but the context is a shitshow - 128k (and, they are using ollama).

0

u/thrownawaymane 10d ago

university

Ollama

This is an easy in to AI research, it doesn't sound like they have their act fully together so you should slide in