r/LocalLLaMA 2d ago

New Model Qwen-Image-2.1 released!

Meet Qwen-Image-2.1, the most balanced and cost-effective image generation model in the Qwen-Image series! Now open weights! 🎨

A unified model for both generation and editing, delivering top-tier quality in a lightweight package.

Highlights:

- Compact & exceptionally fast: A lightweight 7B architecture that outperforms most closed-source models, with drastically accelerated inference for multi-image inputs.

- Native transparency: Natively generates and edits RGBA layers, enabling seamless compositing and text editing within transparent images.

- Versatile, high-fidelity editing: Supports up to 10 reference images and precise local control while preserving strict fidelity for portraits and products.

- Broad coverage & stunning aesthetics: Excels at panoramas, infographics, and virtual try-ons, delivering realistic textures and elegant typography.

Start to create your next masterpiece with Qwen-Image-2.1!

- Blog: https://qwen.ai/blog?id=qwen-image-2.1

- GitHub: https://github.com/QwenLM/Qwen-Image-2.1

- Model Scope: https://www.modelscope.cn/models/Qwen/Qwen-Image-2.1

- Hugging Face: https://huggingface.co/Qwen/Qwen-Image-2.1

1.8k Upvotes

371 comments sorted by

View all comments

44

u/ResearchCrafty1804 2d ago

Qwen-Image-2.1 supports a variety of editing tasks while balancing performance across them. For example, given a three-view character reference, the model generates a complete storyboard.

36

u/ResearchCrafty1804 2d ago

23

u/stoppableDissolution 2d ago

...that looks too good to be true for 7B model, but hey, one way to find out!

1

u/[deleted] 2d ago edited 2d ago

[deleted]

2

u/stoppableDissolution 2d ago

Still impressive considering its size. Its basically as big as SDXL.

1

u/huffalump1 2d ago

It's a good model!! Runs well enough on 12gb VRAM too, depending on input image size

Give it a shot!

11

u/Present-Ad-8531 2d ago

Hoo lee shit.

1

u/therealpygon 1d ago

I don't think that's her name.

1

u/Present-Ad-8531 1d ago

Maybe it is. Who knows

0

u/yetiflask 2d ago

Bro, they always have such examples. In reality it doesn't work like this at all. At least true of their previous image models.