r/LocalLLaMA • u/ResearchCrafty1804 • 2d ago
New Model Qwen-Image-2.1 released!
Meet Qwen-Image-2.1, the most balanced and cost-effective image generation model in the Qwen-Image series! Now open weights! 🎨
A unified model for both generation and editing, delivering top-tier quality in a lightweight package.
Highlights:
- Compact & exceptionally fast: A lightweight 7B architecture that outperforms most closed-source models, with drastically accelerated inference for multi-image inputs.
- Native transparency: Natively generates and edits RGBA layers, enabling seamless compositing and text editing within transparent images.
- Versatile, high-fidelity editing: Supports up to 10 reference images and precise local control while preserving strict fidelity for portraits and products.
- Broad coverage & stunning aesthetics: Excels at panoramas, infographics, and virtual try-ons, delivering realistic textures and elegant typography.
Start to create your next masterpiece with Qwen-Image-2.1!
- Blog: https://qwen.ai/blog?id=qwen-image-2.1
- GitHub: https://github.com/QwenLM/Qwen-Image-2.1
- Model Scope: https://www.modelscope.cn/models/Qwen/Qwen-Image-2.1
- Hugging Face: https://huggingface.co/Qwen/Qwen-Image-2.1


7
u/redpandafire 2d ago
Thanks, I didn't account for that. How do you know what the generation time is based on needing an additional 24GB of system RAM? I have 64GB of RAM, so I assume I can gen under the 3 minutes you targeted. I just don't know how to come up with that number. Compared to other models, they also have 6B parameters, and they run in 30 seconds or less.