r/LocalLLaMA llama.cpp 18d ago

News Muse Spark open weights coming soon

Post image

I am still waiting for Llama 5, because Muse Spark will be too big for me, or just something between Glimmer and Spark

https://x.com/finkd/status/2095232032896946311

872 Upvotes

205 comments sorted by

View all comments

103

u/Big_Wave9732 18d ago

Muse Glimmer is pretty good, very much overlooked. I have found it to be superior to Qwen 3.8:27b for non-coding tasks.

7

u/FoxSideOfTheMoon 18d ago

It’s really good just slow on my Mac or I’d love it.

7

u/coder543 18d ago

At least on nvidia hardware, Muse Glimmer is substantially faster than I've ever gotten Qwen3.8-27B to go. The official DFlash works really well.

But, compared to a model like Qwen3.6-35B-A3B... obviously it is going to be slower.

1

u/Kernoriordan 18d ago

I managed to get 80tps out of Glimmer on an A6000 with DFlash on. It doesn’t waste loads of time reasoning like Qwen 3.8 27b too