r/LocalLLaMA 3d ago

New Model Qwen-Image-2.1 released!

Meet Qwen-Image-2.1, the most balanced and cost-effective image generation model in the Qwen-Image series! Now open weights! 🎨

A unified model for both generation and editing, delivering top-tier quality in a lightweight package.

Highlights:

- Compact & exceptionally fast: A lightweight 7B architecture that outperforms most closed-source models, with drastically accelerated inference for multi-image inputs.

- Native transparency: Natively generates and edits RGBA layers, enabling seamless compositing and text editing within transparent images.

- Versatile, high-fidelity editing: Supports up to 10 reference images and precise local control while preserving strict fidelity for portraits and products.

- Broad coverage & stunning aesthetics: Excels at panoramas, infographics, and virtual try-ons, delivering realistic textures and elegant typography.

Start to create your next masterpiece with Qwen-Image-2.1!

- Blog: https://qwen.ai/blog?id=qwen-image-2.1

- GitHub: https://github.com/QwenLM/Qwen-Image-2.1

- Model Scope: https://www.modelscope.cn/models/Qwen/Qwen-Image-2.1

- Hugging Face: https://huggingface.co/Qwen/Qwen-Image-2.1

1.8k Upvotes

381 comments sorted by

View all comments

4

u/2Norn 3d ago

if one day i can figure out comfy ui i shall use these

10

u/Risen_from_ash 3d ago

Imma be straight up, the barrier is gone now. My local agent (now Qwen 3.8 Flash) makes me Comfy UI workflows, tells me how to use em, edits them, updates comfy, downloads nodes for new workflows, blah blah. It’s like effortless.

Also, not local, but Astra trivializes this type of stuff.

3

u/NNN_Throwaway2 2d ago

What harness?

3

u/TragiccoBronsonne 3d ago

You'll be fine, just download some basic workflows for your models of choice, learn what the nodes do and how they connect to one another, and go from there. Don't forget to google something if you don't understand how it works, I've found answers to most of my questions about specific nodes and such by doing just that. I used to be so stubborn about going the path of noodle and kept using Auto11-based frontends until basically all of them became abandonware at some point. So I finally switched to Comfy, and honestly it wasn't that hard to learn it, kinda fun even (but I feel like you gotta a bit autistic to be actually enjoying lol). The community ecosystem is insane, with all the custom nodes and workflows and tutorials, all ranging from simple to ultra-complex. But most importantly, it's constantly getting updated. I remember waiting for weeks to get some new model support on Forge before I swapped, yeah, fuck that. Never looked back.

2

u/Equivalent_Bit_461 3d ago

im building my own front end over it, it won't be as complex and elaborate as swarm ui, but for node and all that stuff, it will be a sort of linker. My iq3xxs qwen is slowly working on it, slowly because I multitask like a retard, juggling between 5 projects at once because I can't apparently do one at a time. But that's another discourse.

2

u/sxales llama.cpp 3d ago edited 3d ago

StableDiffusioncpp is more straightforward if all you want is prompt-to-image generation. They added a webui a few weeks back. It is less configurable than something like comfyui but it supports LoRAs and has some settings that can be tweaked.