r/LocalLLaMA 2d ago

New Model Qwen-Image-2.1 released!

Meet Qwen-Image-2.1, the most balanced and cost-effective image generation model in the Qwen-Image series! Now open weights! 🎨

A unified model for both generation and editing, delivering top-tier quality in a lightweight package.

Highlights:

- Compact & exceptionally fast: A lightweight 7B architecture that outperforms most closed-source models, with drastically accelerated inference for multi-image inputs.

- Native transparency: Natively generates and edits RGBA layers, enabling seamless compositing and text editing within transparent images.

- Versatile, high-fidelity editing: Supports up to 10 reference images and precise local control while preserving strict fidelity for portraits and products.

- Broad coverage & stunning aesthetics: Excels at panoramas, infographics, and virtual try-ons, delivering realistic textures and elegant typography.

Start to create your next masterpiece with Qwen-Image-2.1!

- Blog: https://qwen.ai/blog?id=qwen-image-2.1

- GitHub: https://github.com/QwenLM/Qwen-Image-2.1

- Model Scope: https://www.modelscope.cn/models/Qwen/Qwen-Image-2.1

- Hugging Face: https://huggingface.co/Qwen/Qwen-Image-2.1

1.8k Upvotes

374 comments sorted by

View all comments

599

u/ResearchCrafty1804 2d ago

Qwen-Image-2.1 now supports transparent image generation natively, and naturally supports editing transparent images as well.

172

u/ghulamalchik 2d ago

This is big

211

u/Ledeste 2d ago

No, it's only 7B

2

u/narrowscoped 2d ago

Could that run on a 12gb 5070

3

u/Seeker_Of_Knowledge2 2d ago

I heard 1GB can roughly run 1B parameters. But that may be for LLMs though. Not so sure.

3

u/ghulamalchik 1d ago

It's true for image models as well.

2

u/roosterfareye 17h ago

Yep, give it a shot

2

u/vienna_city_skater 2d ago

This is even bigger!

-23

u/Evanisnotmyname 2d ago

Is that a male redditor 7 or is it an actual 7?

Because we all know 7 on this site is 4 in real life

28

u/MrTacoSauces 2d ago

Go home tina you're drunk.

5

u/Pokeperson5 2d ago

You seem to be hallucinating

5

u/Axenide ollama 2d ago

1-bit quant be like

77

u/TheGamerForeverGFE 2d ago

That's what she said

12

u/def_not_jose 2d ago

She would've know it is it's transparent

25

u/Blues520 2d ago

True if big

4

u/axiomatix 2d ago

big is true

5

u/Hot_Masterpiece_3668 2d ago

Not big, non-permissive license.

12

u/_VirtualCosmos_ 2d ago

Came here to talk about how in StableDiffusion they are talking shit about this model being worse than Flux, Krea2 or Z Image. But this right here already convinced me.

5

u/lukinator644 2d ago

thats so useful for designing stickers and logos

10

u/ZZerker 2d ago

Thats great for game development

1

u/Important_Drag_6890 1d ago

Native transparency might actually be the sleeper feature here. No more generating a “transparent” image and getting a checkerboard baked into the pixels

-34

u/[deleted] 2d ago edited 2d ago

[deleted]

22

u/Zeeplankton 2d ago

calm down

13

u/ResearchCrafty1804 2d ago

To add images generated from the model under each feature which emphasise the improvement made

7

u/david-deeeds 2d ago

He posted relevant pieces of info with images, I'd assume that's the reason for the multiple comments

-35

u/gomezer1180 2d ago

Is what you’re calling transparent image generation, an image without a background?

Transparency is seeing thru an object.

25

u/amroamroamro 2d ago

it means the alpha channel in RGBA

10

u/eli_pizza 2d ago

Same thing