50% OFF Your first purchase is half price — the discount is already in the price you see. Ends September 30.

Qwen Image 2.1 AI Image Generator

Qwen Image 2.1 in your browser: turn a prompt or several photos into a finished image, with no graphics card and no ComfyUI.

Nothing to installSeveral photos, one imageFailed runs refunded

Qwen Image 2.1 AI Image Generator

15 creditsSign in to run this one
Examples

Qwen Image 2.1 images people made

Shared by the people who made them, plus one made here on Kavel.

A flooded Tokyo street, every detail kept
A studio mug, dropped into a street at sunset
A cottage floating over a misty field

A flooded Tokyo street, every detail kept

How we made it

Put a product into a scene with Qwen Image 2.1

The street and mug example above, in four moves.

01

Upload the scene, then the product

Order matters: the first photo is the scene, the second is the object going into it.

02

Name each photo by its place

Write "the mug from the second image" and say exactly where it goes in the first.

03

Say what must not move

List what stays: the street, the window, the words on the board.

04

Fix one detail in a second pass

Upload the result and change a single word or object, at 2K for the keeper.

Access

Qwen Image 2.1 is open weights. You still don't need the GPU

Alibaba put the weights on Hugging Face in September 2026, and people run them on a 16GB graphics card through ComfyUI. Kavel runs the same model for you, in the browser.

Qwen Image 2.1 — a sunlit coral reef full of fish
A sunlit coral reef from @ItsmeAjayKV, rendered on a home RTX 3090. Here the same model runs on our hardware instead.

Creative engine

Running on Kavel

2K
Top resolution
7 ratios
Square to widescreen
15 credits
Per 1K image, edits too

A prompt or photos in

Type a scene, or upload up to nine photos and say how they combine.

Short prompts expanded

Every run rewrites your prompt into a fuller scene description before it draws.

The price before you press

The exact cost shows before you generate, and a run that fails on our side is refunded.

What people post about

Three jobs Qwen Image 2.1 does best, going by what people post

Furnish a room from nine product photos

Furnish a room from nine product photos

"I used 9 image references to furnish the room, using 9 real furniture pieces to design the workspace," @junwatu wrote beside this render. On r/StableDiffusion, the thread calling Qwen Image 2.1 an editing beast drew a hundred replies.

  • Upload the room first, then each piece
  • Refer to them as the first image, the second image
  • Keep it to four photos when detail matters most
Impossible scenes that look photographed

Impossible scenes that look photographed

@superalesha wanted "ordinary places where something impossible happens and people just carry on with their day" and posted the prompts. The paper boat on a canal is one of that set.

  • Describe the ordinary place first
  • Add the one impossible thing, with its size
  • Ask for documentary photography, not fantasy
Anime and illustration, not just photos

Anime and illustration, not just photos

This riverside scene is a plain text-to-image run posted by @PSueoka55133: evening light on the water, a bike on the railing, a train on the bridge.

  • Name the medium: anime illustration, watercolour
  • Set the hour and the light
  • Square or 2:3 for a single character
Prompts

Qwen Image 2.1 prompts you can copy

Two prompts people posted beside the image they got, word for word, and the one we used for the mug above.

The found-footage convenience store

The found-footage convenience store

@cocktailpeanut

Camera type and lighting first, then the one thing that is wrong in the frame.

A grainy surveillance-like found-footage frame inside a nearly empty convenience store late at night. The fluorescent lighting is flat and ugly. Between the aisles, a person-shaped figure with an unnaturally featureless face is caught mid-turn looking toward the camera. Shelves of snacks and drinks make the scene feel totally ordinary, which makes the figure more disturbing. Slight distortion, low dynamic range, timestamp, accidental realism, not stylized.
Run this prompt
A studio product shot

A studio product shot

Kavel

Material, surface, backdrop and one light, in that order. We used it as the second photo in the street edit.

Studio product photo of a speckled sage-green ceramic coffee mug with a thick hand-thrown rim and a curved handle, standing on a pale oak board against a soft warm-grey backdrop, single soft window light from the left, gentle shadow, sharp glaze texture.
Run this prompt
What people say

What people said after seeing Qwen Image 2.1 results

Reactions from the first Reddit threads, quoted as written.

“

Is this AI?

U
r/StableDiffusion
Compared

How Qwen Image 2.1 compares with Nano Banana 2, Qwen Image 3, Flux and a local install

Matchups people ran themselves, plus what changes when you run it here.

vs Qwen Image 3

Same family, different jobs. Qwen Image 3 Pro is the one built for pages full of text, with no weights to download. Version 2.1 is the open-weights model for editing with several photos, at three fifths of the credits.

See Qwen Image 3

vs Flux

A reply in the license thread on r/LocalLLaMA, from someone already running it.

“Truly is a phenomenal model and easily beats all current Flux models .”

vs running it on your own card

An 8GB RTX 4070 owner on r/comfyui could not keep it running. Here there is no card, no driver and no workflow to break.

“When I run Qwen Image 2.1 from time to time it crashes with error message below.”

FAQ

Qwen Image 2.1 FAQ

The questions people ask most in the Reddit and X threads about it.

The weights are public on Hugging Face under Qwen's own license. You don't need them here: pick the model, type or upload, and generate in the browser.

None on Kavel, because it runs on our hardware. Locally, people size it for a 16GB graphics card, and one 8GB RTX 4070 owner reported ComfyUI crashing on every run.

You can start free: log in and Kavel adds free credits with no card, which run Kavel's own image engine. This model runs on any paid plan or credit pack, with the cost shown before you generate.

15 credits at 1K and 30 at 2K, edits included. Failed runs are refunded automatically, unless the prompt broke the content policy.

Upload up to nine photos in one run here. When a face or a product has to stay exact, give it four references or fewer and name what must not change.

By upload order. The first photo you add is "the first image" in your prompt, the second is "the second image", and so on.

Not here. Locally the official rewriter is a separate model of its own. On Kavel every run rewrites your prompt into a fuller scene before drawing, so a short prompt still gets a full picture.

Yes. Put each line in quotation marks and say where it sits, as in the flooded Tokyo example with its window sign. Render the keeper at 2K.

Make your first Qwen Image 2.1 image

Skip the 16GB card and the ComfyUI install. Type a prompt or drop in your photos, pick 1K for 15 credits or 2K for 30, and leave with the image. A run that errors out on our side gives its credits back.