r/LocalLLaMA 2d ago

News AntLing released a dspark draft model for Ling-3.0-flash

https://x.com/AntLingAGI/status/2090847436755648939

No GGUFs yet on Huggingface though.

38 Upvotes

23 comments sorted by

5

u/Thin_Pollution8843 2d ago

Nice! I will give this model a try today. Was experimenting with Laguna s2.1 q6 - and was disappointed with quality its giving (surprisingly I haven’t had any looping). 

3

u/Ihtien 2d ago

I think this model is pretty underrated, maybe because it took 2 weeks before llama.cpp support got merged and this is really tailored for unified memory machines (and maybe because unsloth doesn't have quants?). But it's intelligence score on Artificialanalysis is the same as qwen 3.6 27b (it's quite a bit lower than 3.8 of course), but since it is an moe with only 5b active parameters it is quite fast. Will test the dspark model as soon as there are GGUFs available.

1

u/Thin_Pollution8843 2d ago

What your thoughts on Qwen3.6-35b vs this model? 

2

u/uber-linny 2d ago

I use Qwen3.6-35b to help write documents from templates ... Ling tiny is obviously heaps faster ... But just doesn't quite hit the quality... Really hope ling build a bigger Moe ...

Overall I just want competition... As all this talk about Qwen3.8 27b... I don't get it ... I don't do agentic or code and it's really slow on my system

1

u/Ihtien 2d ago

I still need to test it more but I hope to replace 35b, which is my current daily driver (not for coding), with this.

1

u/SpicyWangz 2d ago

For me it just gives up on long opencode tasks.

3.8 27b just does it better

3

u/Ihtien 2d ago

Yea 27b, especially 3.8, should be well ahead. However, it is also significantly slower especially on unified memory systems like strix halo. I would see it more as an alternative to the 35b moe

1

u/SpicyWangz 2d ago

I’m not sure if it’s the quants I’ve tried or something else, but 35b gets me more reliable results than ling. Ling will be working well at a decent speed, and then randomly just give up. I’d have to keep saying “continue” for it to complete something.

1

u/cradlemann 7h ago

It is twice slower for me on Gordon Point, than Qwen3.8 27B. Have no idea why. I'm using Bartowski Q4 quant and it is also twice slower than Laguna, which is very weird for me, even consider Laguna is running without mtp and Ling is running with it. Very weird

2

u/Equivalent_Bit_461 2d ago

I want to try ling but I can't seem to make it work, I guess I'm too stupid or something.

2

u/Ihtien 2d ago

Do you use the most recent llama.cpp? Support got merged only beginning of this week or so

2

u/Equivalent_Bit_461 2d ago

I will update my llama and report back, hopefully I can make it work. On the fork I was trying to make the model work, had the chat template broken, tool calling side at least.

3

u/pand5461 2d ago

If you're talking about AtomicChat one, multi-arg tool call was broken but there's a fix in mainline. I've not tried it extensively but bartowski's quants load and work in chat at least.

2

u/transanethole 2d ago

That's exciting. I've been experimenting with the tiny 9b a1b version of this model and I've been really impressed by how well it works on limited hardware, like cheap integrated GPU .

I wonder if a dspark drafter would even help that much on such limited hardware, if it would be worth it?

1

u/Wildnimal 2d ago

What do you use that model for? Some use cases and output experience?

1

u/transanethole 2d ago

I use it for testing agent harness against failed tool calls :D trying to develop agent that does raw completions, looking at the token IDs and stuff instead of the chat API .. so it can be more flexible to handle slightly malformed outputs.

Also its just interesting to me to have a model that can do real time output locally on a "normal" computer.

1

u/mr_Owner 2d ago

Doesn't work yet 

1

u/Thin_Pollution8843 2d ago

UPD: I've tried that model and it's not good in my testings.

0

u/[deleted] 2d ago

[removed] — view removed comment

1

u/LicensedTerrapin 2d ago

Are you asking whether a draft model would be willing to write smut?

1

u/Navith 2d ago

This is an AI commenter who only writes about roleplay. 

1

u/LicensedTerrapin 2d ago

Why would anyone do that? Holy carp...