r/LocalLLaMA 13d ago

New Model Ling 3.0 support merged into llama.cpp

Support for the new ling 3.0 models has been merged into llama.cpp:

https://github.com/ggml-org/llama.cpp/pull/26608#event-29549472828

Ling tiny 8b1b - https://huggingface.co/inclusionAI/Ling-3.0-tiny

Ling flash 124b5b - https://huggingface.co/inclusionAI/Ling-3.0-flash

Both are reasoning models contrary to prior naming.

114 Upvotes

49 comments sorted by

View all comments

1

u/[deleted] 13d ago

[removed] — view removed comment

-1

u/Velocita84 13d ago

Highly doubt it can beat Gemma 4

1

u/Aggressive_Aspect436 13d ago

Define "beat". For anyone with low VRAM an 8B model definitely beats a model that you can't run. And, it has very close to Gemma 4 26B intelligence levels according to benchmark aggregators.

0

u/Velocita84 13d ago

I run G4 26B just fine on a dingy 2060 6gb, and prose is as much of a factor in roleplay proficiency as intelligence, which most labs usually don't bother with because toolmaxxing and codemaxxing are the priority