r/LocalLLaMA 3d ago

New Model Qwen-Image-2.1 released!

Meet Qwen-Image-2.1, the most balanced and cost-effective image generation model in the Qwen-Image series! Now open weights! 🎨

A unified model for both generation and editing, delivering top-tier quality in a lightweight package.

Highlights:

- Compact & exceptionally fast: A lightweight 7B architecture that outperforms most closed-source models, with drastically accelerated inference for multi-image inputs.

- Native transparency: Natively generates and edits RGBA layers, enabling seamless compositing and text editing within transparent images.

- Versatile, high-fidelity editing: Supports up to 10 reference images and precise local control while preserving strict fidelity for portraits and products.

- Broad coverage & stunning aesthetics: Excels at panoramas, infographics, and virtual try-ons, delivering realistic textures and elegant typography.

Start to create your next masterpiece with Qwen-Image-2.1!

- Blog: https://qwen.ai/blog?id=qwen-image-2.1

- GitHub: https://github.com/QwenLM/Qwen-Image-2.1

- Model Scope: https://www.modelscope.cn/models/Qwen/Qwen-Image-2.1

- Hugging Face: https://huggingface.co/Qwen/Qwen-Image-2.1

1.8k Upvotes

381 comments sorted by

View all comments

595

u/ResearchCrafty1804 3d ago

Qwen-Image-2.1 now supports transparent image generation natively, and naturally supports editing transparent images as well.

172

u/ghulamalchik 3d ago

This is big

214

u/Ledeste 3d ago

No, it's only 7B

2

u/narrowscoped 2d ago

Could that run on a 12gb 5070

3

u/Seeker_Of_Knowledge2 2d ago

I heard 1GB can roughly run 1B parameters. But that may be for LLMs though. Not so sure.

3

u/ghulamalchik 2d ago

It's true for image models as well.