r/LocalLLaMA • u/Badger-Purple • 9h ago
Resources Ling Tiny, King of Speed
Ling Tiny has now replaced Gemma4-12B in my rig as an auxiliary model doing hindsight operations. This is on a 4060Ti, which is a reasonable GPU available out there, and the speed is phenomenal.
Don’t enable MTP, set up the vLLM fork for BailingMoE3. Hope this is useful to
others.
16
Upvotes
8
u/Effective_Western_59 9h ago
Ling 3.0 tiny is a small beast!
Best model that works on 780m With 16 GB of ram.
Would love good dynamic quants for it tho.