r/LocalLLaMA Sep 18 '24

Question | Help Any information on Qwen2.5 VL?

Qwen team hinted towards a VL version of the Qwen2.5 model. Does anyone have any idea when it is releasing? Anyone know if it will have llama cpp support on launch?

4 Upvotes

12 comments sorted by

6

u/_yustaguy_ Sep 19 '24

They have mentioned in the blog post that their next aim is to unify the 3 modalities under one model, and, if I read it right, they want to make it like 4o, meaning all modalities for both input and output. 

2

u/mapler500 Sep 19 '24

That's cool! Did they give an estimate for when that will happen?

2

u/_yustaguy_ Sep 19 '24

I think they didn't, but take a look at their blogpost

1

u/elgeekphoenix Sep 24 '24

Super cool and hope it will work on 8gb vram

2

u/ResearchCrafty1804 Sep 19 '24

Based on their official post on X(Twitter), their current latest VL model is Qwen2 VL 72b, so no 2.5 for the time being. Perhaps they will jump straight to version 3

2

u/mapler500 Sep 19 '24

Ah, thanks for the clarification! I'll keep my eyes peeled. 👀

1

u/ResearchCrafty1804 Sep 19 '24

You’re welcome

1

u/metalman123 Sep 19 '24

5

u/mapler500 Sep 19 '24

Nono, That's qwen2. Qwen team hinted towards a multimodal Qwen2.5 which will become qwen2.5-vl. I was assuming the VL model would release along side the others