logo
RunComfy
  • ComfyUI
  • TrainerNew
  • Models
  • API
  • Pricing
discord logo
MODELS
Explore
All Models
LIBRARY
Generations
MODEL APIS
API Docs
API Keys
ACCOUNT
Usage

Qwen Image 2.1: Text-to-Image and Image Editing on Models and API | RunComfy

qwen/qwen-image-2.1/edit

Qwen Image 2.1 writes stills from text and edits up to ten references, with native transparency, 1K or 2K output, and seed lock.

Image 1
One to ten reference images. Order matches the prompt (the first image, the second image). JPEG, PNG, or WebP; each file up to 30MB and 25MP. Keep to four or fewer when identity must stay sharp.
Describe the change in plain language: what to add, remove, restyle, or replace. Quote any text to render it verbatim. With a mask, describe what should appear in the white area.
Optional black-and-white local-edit mask. White marks the area to change; black stays. Requires exactly one reference image, cannot pair with a transparent background, and ignores aspect ratio and prompt rewrite.
Output aspect ratio. auto uses the first reference. Ignored in mask mode.
Output detail tier. 1K is faster; 2K has about four times the pixels and costs more.
opaque for a regular photo. transparent writes a real alpha channel — describe only the subject, and do not use JPEG. Cannot combine with a mask.
PNG and WebP can carry an alpha channel. JPEG cannot, so it cannot be combined with a transparent background.
Rewrite a short instruction into a fuller scene before editing. Turn off when the instruction is already exact. Ignored in mask mode.
Random seed for reproducibility. Use -1 for a random seed each run.
Idle
The rate is $0.029 per 1K output image and $0.119 per 2K output image.

Introduction To Qwen Image 2.1

Qwen's Qwen Image 2.1 creates new stills from a text prompt and revises photos you upload, with native transparency at 1K or 2K.
Trading separate generators and cutout tools for one compact model, it helps designers, marketers, and product teams ship covers, catalog frames, and edits together.
For developers, Qwen Image 2.1 on RunComfy can be used both in the browser and via an HTTP API, so you don't need to host or scale the model yourself.
Ideal for: Text-to-Image Posters | Reference Image Edits | Transparent Cutouts

Qwen / Qwen Image 2.1#


This compact stills model both writes a new picture from text and revises photos you already have. One set of Qwen Image 2.1 weights covers text-to-image and image-to-image, at 1K or 2K, with optional native transparency.


Prompt-only generation needs no upload. Editing adds one to ten reference images plus an instruction. This page runs the Qwen Image 2.1 reference workflow; a sibling page runs prompt-only generation.


Highlights#


  • Text-to-image: Qwen Image 2.1 turns a written brief into a finished still — posters, product shots, portraits, and layouts with readable type.
  • Image-to-image: The same weights restage a photo from a plain-language instruction: swap a backdrop, restyle a product, or combine several references.
  • Native transparency: Qwen Image 2.1 can return an RGBA sticker, icon, or product cutout, or lift a subject off a regular RGB photo onto a clear layer.
  • Up to ten references: Merge portraits into a group, dress a model from garment photos, or furnish a room from product stills.
  • Local region control: Qwen Image 2.1 can follow a circle, a painted mark, or a black-and-white mask. Mask mode needs exactly one reference and cannot pair with a transparent background.
  • People, products, and type: Faces, packaging, and headlines stay closer to the brief, which helps catalog frames and covers.

Wider jobs also work with Qwen Image 2.1: expanding a selfie into a panorama, turning a product photo into an infographic, or building a storyboard from a three-view character sheet.


Parameters#


On this page, Qwen Image 2.1 runs image-to-image. Text-to-image uses the same resolution, ratio, background, format, prompt rewrite, and seed choices, without reference images or a mask.


ParameterRequiredTypeDefaultRange / OptionsDescription
image_urls*Yes (*)ArraySample image1–10 imagesReference images in prompt order. JPEG, PNG, or WebP; each file up to 30MB and 25MP. Keep to four or fewer when identity must stay sharp.
prompt*Yes (*)StringExample instructionUp to 5,000 charactersDescribe the new still or the edit. Quote any on-image text. With a mask, describe what should appear in the white area.
mask_urlNoString (image)—OptionalBlack-and-white local-edit mask. White changes, black stays. Requires exactly one reference. Aspect ratio and prompt rewrite are ignored. Cannot pair with a transparent background.
aspect_ratioNoStringautoauto, 1:1, 4:3, 3:4, 3:2, 2:3, 16:9, 9:16, 21:9, 9:21auto snaps to the first reference. Ignored in mask mode. Text-to-image defaults to 1:1.
resolutionNoString1K1K, 2K1K is faster; 2K has about four times the pixels and takes longer, especially with references.
backgroundNoStringopaqueopaque, transparenttransparent writes a real alpha channel for new stills or cutouts. Describe only the subject. Cannot pair with JPEG.
output_formatNoStringpngpng, webp, jpegPNG and WebP can carry alpha. JPEG cannot.
enhance_promptNoBooleantruetrue, falseRewrites a short prompt before generating or editing. Turn off when it is already exact. Ignored in mask mode.
seedNoInteger-1-1 or 0–2147483647Fix a seed to reproduce a result; use -1 for a new variation.

  • Required field.

Pricing#


On this image-to-image page, Qwen Image 2.1 is $0.029 per 1K image and $0.119 per 2K image. Text-to-image is $0.020 per 1K image and $0.079 per 2K image on its own page. Extra references do not add a listed surcharge. For a batch, multiply the selected-tier rate by the number of outputs.


How to Use#


  1. Use Qwen Image 2.1 for text-to-image when you only have a prompt, or image-to-image when you have photos to revise.
  2. For a new still, name the subject, setting, style, and any exact wording in quotes.
  3. For an edit, upload one to ten references in the order the prompt names them.
  4. Say what Qwen Image 2.1 should change and what should stay. Add a mask only when one region should move, and use a single reference in that mode.
  5. Choose 1K for a faster Qwen Image 2.1 draft or 2K for delivery detail.
  6. Pick an aspect ratio, or leave auto on an edit so the first reference sets the frame.
  7. Ask Qwen Image 2.1 for a transparent background when you need a cutout. That output needs PNG or WebP, not JPEG.
  8. Generate, then change one instruction at a time. Fix a seed when you want the same look again.

Prompt & Reference Tips#


  • For a new still, lead with the subject, then lighting, lens, and where Qwen Image 2.1 should place any headline.
  • Put on-image wording in quotes so it is painted verbatim.
  • Name references in order: "put the person from the first image into the room from the second image."
  • Keep the Qwen Image 2.1 reference count at four or fewer when identity or logos must stay sharp.
  • For group shots or try-on, a landscape ratio such as 3:2 or 16:9 usually fits more subjects.
  • For transparent cutouts, describe only the subject. Mentioning a scene often makes Qwen Image 2.1 fill the background back in.
  • When using a mask, describe what should appear in the white area.
  • Iterate on a fixed seed: change one of backdrop, wardrobe, or lighting per run.

Qwen Image 2.1 accepts mixed Chinese and English prompts. Short, concrete nouns beat long mood boards.


How Qwen Image 2.1 compares to other models#


  • Based on publicly available information, it stands out among compact open stills models because generation, editing, and native RGBA output share one set of weights.
  • Against prompt-only generators, it also accepts up to ten references and can target a region with a mask, circle, or painted mark.
  • Dedicated cutout tools only remove backgrounds. Qwen Image 2.1 can create a transparent asset from text or extract one from a photo.
  • For a single photoreal portrait with no type and no compositing, a larger closed generator may still win on skin micro-detail.

More Models to Try#


If Qwen Image 2.1 is not the right starting point, compare these models on RunComfy:


  • Qwen Image 3.0 — later Qwen text-to-image stills.
  • Qwen Image 3.0 Edit — instruction edits on the Qwen Image 3.0 weights.
  • Qwen Edit 2509 — multi-image edits with pose and layout locks.
  • Nano Banana Pro — detailed text-to-image stills.
  • Flux 2 Pro — studio-style prompt-only stills.

Official Resources#


  • Qwen: https://qwen.ai/
  • Qwen blog — Qwen Image 2.1: https://qwen.ai/blog?id=qwen-image-2.1
  • Hugging Face weights: https://huggingface.co/Qwen/Qwen-Image-2.1
  • GitHub: https://github.com/QwenLM/Qwen-Image-2.1

Related Models

flux-2/pro/edit

Edit detailed visuals fast with layout-aware, multi-reference control for brand-ready results.

ovis-image

Produce high-fidelity visuals with clear text, fast generation, and professional design control.

qwen-image/qwen-image-edit-2511

Advanced image-to-image tool with geometry-aware edits and consistent identity control for creative workflows.

flux-2/pro/text-to-image

Create reliable, studio-grade visuals with precise color and layout control.

qwen-edit-2509/lora/fusion

Blend and refine visuals with advanced image editing, depth control, and multilingual design precision.

qwen-image-3.0/pro/edit

Qwen Image 3.0 Pro Edit: pro instruction-based image editing

Frequently Asked Questions

What is Qwen Image 2.1 used for in image-to-image workflows?

Qwen Image 2.1 revises photos from a written instruction while keeping identity, product shape, and layout. Typical jobs include background swaps, virtual try-on, group composites, local inpainting, and transparent product cutouts.

How many reference images can Qwen Image 2.1 take?

You can upload one to ten reference images. Qwen Image 2.1 reads them in array order, so "the first image" and "the second image" in the prompt map to that list. Keep the count at four or fewer when faces or logos must stay sharp.

Does Qwen Image 2.1 support transparent backgrounds and subject cutouts?

Yes. Set background to transparent so Qwen Image 2.1 writes a real alpha channel, or ask it to lift a subject from a regular RGB photo onto a clear layer. Use PNG or WebP; JPEG cannot carry alpha, and a mask cannot be combined with a transparent background.

How do local edits and masks work in Qwen Image 2.1?

Qwen Image 2.1 can follow circles, painted marks, or a separate black-and-white mask (white changes, black stays). Mask mode needs exactly one reference, follows that image's ratio, and ignores aspect ratio and prompt rewrite. Describe what should appear in the white area.

What output sizes does Qwen Image 2.1 support?

Qwen Image 2.1 offers 1K or 2K output and aspect ratios including auto, 1:1, 4:3, 3:4, 3:2, 2:3, 16:9, 9:16, 21:9, and 9:21. auto snaps to the first reference. Check the current RunComfy parameter panel for the exact options.

Can Qwen Image 2.1 generate images from text as well as edit them?

Yes. The same Qwen Image 2.1 weights handle text-to-image and image-to-image. This page is the Edit / reference workflow; a separate text-to-image page generates from a prompt without uploading photos.

Can developers use Qwen Image 2.1 through the RunComfy API?

Yes. Prototype an edit in the RunComfy model UI, then call the same Qwen Image 2.1 model via the RunComfy HTTP API with identical parameters. You do not need to host or scale the model yourself.

How much does Qwen Image 2.1 cost on RunComfy?

Generations consume usd or credits. Qwen Image 2.1 is billed at $0.029 per 1K output image and $0.119 per 2K output image. Extra references do not add a listed surcharge on this page; new users typically get a free trial amount to test with.

Follow us
  • LinkedIn
  • Facebook
  • Instagram
  • Twitter
Support
  • Discord
  • Email
  • System Status
  • Affiliate
Video Models
  • Runway Aleph 2
  • Wan 2.5
  • Wan 2.6 Flash
  • MiniMax H3 Max
  • Wan 3.0 Reference To Video
  • Seedance 2.5 Reference to Video 480p
  • View All Models →
Image Models
  • Wan 2.6 Image to Image
  • Seedream 5.0 Pro
  • Flux 2 Klein 9B
  • Nano Banana Pro
  • Nano Banana 2 Edit
  • Z Image Turbo LoRA
  • View All Models →
Legal
  • Terms of Service
  • Privacy Policy
  • Cookie Policy
RunComfy
Copyright 2026 RunComfy. All Rights Reserved.

RunComfy is the premier ComfyUI platform, offering ComfyUI online environment and services, along with ComfyUI workflows featuring stunning visuals. RunComfy also provides AI Models, enabling artists to harness the latest AI tools to create incredible art.

Examples Of Qwen Image 2.1