cd /news/generative-ai/qwen-image-2-1-in-comfyui-open-weigh… · home › topics › generative-ai › article
[ARTICLE · art-138364] src=blog.comfy.org ↗ pub= topic=generative-ai verified=true sentiment=↑ positive

Qwen-Image-2.1 in ComfyUI: Open-Weight Image Generation and Editing, Now with Transparency

Alibaba's Qwen team released Qwen-Image-2.1, a 7B-parameter open-weight image generation and editing model that is now natively supported in ComfyUI, with weights published on Hugging Face. The model runs on an optimized MMDiT architecture, outputs native 2048×2048 images with a real RGBA alpha channel, and accepts up to 10 reference images through the Text Encode Qwen Image 2.1 node. It follows Qwen-Image 2.0, which shipped in February 2026 with native 2K output and transparency.

by read1 min views6 publishedSep 20, 2026
Qwen-Image-2.1 in ComfyUI: Open-Weight Image Generation and Editing, Now with Transparency
Image: Blog (auto-discovered)

A new open-source image editing model has finally arrived. Qwen-Image-2.1 is now supported natively in ComfyUI. Open weights, 7B, and it generates a real alpha channel.

Qwen-Image 2.1 is the latest image model from Alibaba’s Qwen team. It runs 7B parameters on an optimized MMDiT architecture — the same class as Qwen-Image 2.0, which shipped in February 2026 with native 2K output, images with transparency, high quality text rendering and editing capabilities that takes up to 10 input images at once At 7B, inference is fast and the weights fit comfortably on consumer cards.

Model Highlights

  • RGBA output. Four channels, alpha included. Sprites, logos, icons, and product cutouts come out of the sampler ready to composite. No background removal node, no matting model, no edge cleanup. No other major open model does this.
  • Native 2K generation. 2048×2048 direct output. Generated at that resolution, not upscaled into it.
  • Up to 10 reference images. The Text Encode Qwen Image 2.1 node opens image inputs as you fill them, to 10. Character, product, background plate, style reference — all read by the text encoder and spliced into the sequence as VAE latents.
  • One checkpoint. Generation and editing in the same model.
  • 7B parameters. Fast inference, low cost, consumer VRAM.

Getting Started

1. Update ComfyUI to the latest version, or open [Comfy Cloud](https://cloud.comfy.org/) .
2. Download the Qwen-Image-2.1 weights from [Hugging Face](https://huggingface.co/Comfy-Org/Qwen-Image-2.1) and drop them in your models folder.
  1. Load the Qwen-Image-2.1 template from the Templates panel, or download the workflow here .
  2. Write your prompt, attach reference images to image_1 onward, and run

Try it on Comfy Cloud, and let us know what you make. As always, enjoy creating!

── more in #generative-ai 4 stories · sorted by recency
── more on @qwen-image-2.1 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
→ Live at https://your-agent.zahid.host ✓
Get free account → Pricing
from €0/mo · no card required
LIVE [news/qwen-image-2-1-in-co…] indexed:0 read:1min 2026-09-20 · —