50% OFF Your first purchase is half price — the discount is already in the price you see. Ends September 30.

Qwen Image 3 AI Image Generator

Qwen Image 3 turns a written brief into infographics, storyboards and photos where every label reads cleanly, right in your browser with nothing to download.

Words you can readEdit an image you uploadFailed runs refunded

Qwen Image 3 AI Image Generator

25 creditsSign in to run this one
Examples

Qwen Image 3 images people posted

Posted by the people who made them, model named. Their prompts run here in one click.

A five-panel timeline, every label in place
A whiteboard in Japanese and English
An old Showa coffee shop

A five-panel timeline, every label in place

How creators work

Turn a brief into an infographic

The five-panel timeline test on r/AIGenArt, in four moves.

01

Name the layout and the title

Say the format, a wide banner with five equal panels, then quote the headline.

02

Fill each panel in order

A date, a name and three short bullet lines for each panel, every word spelled out.

03

Set the style and the ratio

Colours, type hierarchy and sharp readable small text, at 16:9.

04

Render at 2K and zoom in

45 credits buys the size where captions stay crisp.

Access

Qwen Image 3 has no weights to download. Run it here instead

Alibaba has published no weights as of September 2026, and ComfyUI supports it through API nodes only. Kavel runs the Pro tier for you, in the browser.

Qwen Image 3.0 Pro — a festival poster with a fully legible headline
A risograph poster reading NORTHERN LIGHT FILM FESTIVAL, date and venue intact. Generated on Kavel with Qwen Image 3.0 Pro.

Creative engine

Running on Kavel

2K
Top resolution
5 ratios
Square, wide, vertical
25 credits
Per 1K image, edits too

A brief or an image in

Write the layout block by block, or upload a picture and say what changes.

1K or 2K out

Draft at 1K, render the keeper at 2K.

The price before you press

The exact cost shows before you generate, and a run that fails on our side is refunded.

What people post about

Three jobs Qwen Image 3 handles best, according to Reddit

A whole page of information in one image

A whole page of information in one image

"Every date, model name, and bullet point matched the brief exactly," a designer wrote after an infographic test on r/AIGenArt. This cold brew guide took one run here.

  • One sentence per section, in reading order
  • Exact words in quotation marks
  • Final render at 2K for 45 credits
Skin and fabric that hold up when you zoom

Skin and fabric that hold up when you zoom

The same tester found "genuine pore-level texture and natural variation rather than airbrushed smoothness" in a close-up. Here: warm skin, a patterned sarong, rattan weave.

  • Name each texture: pores, brow hairs, woven linen
  • Give light and lens: window light, 100mm at f/4
  • Judge the detail at 2K, zoomed in
Storyboards and key frames for a video

Storyboards and key frames for a video

A 90-second short on r/comfyui started as six Qwen Image 3 Pro key frames. This sheet, made here, holds six numbered shots.

  • One shot per panel: framing, subject, action
  • A caption line under every panel
  • Animate a frame in MiniMax H3
Prompts

Qwen Image 3 prompts you can copy

Long, specific prompts are where this model shines. Here are two that people posted beside the image they got, word for word.

A café portrait with a wall script

A café portrait with a wall script

@woleswoosh

Outfit, pose, then the room, with the wall words in quotation marks.

Photorealistic full-body portrait of a young East Asian woman (Myanmar features) with long, wavy black hair parted with soft bangs framing her face, fair skin, subtle natural makeup, and soft red lipstick. She has a gentle, slightly playful expression while looking directly at the camera.

She is wearing a tight, ruched black floral mini dress covered in small pink, yellow, and white flowers, with a square neckline and a flared ruffle hem that ends mid-thigh. Over it she wears a sheer, translucent bright red long-sleeve bolero/shrug with dramatic flared bell cuffs.

She stands casually leaning her right hip against a round wooden stool, left hand resting lightly on the stool edge, right arm relaxed at her side.

Setting: modern minimalist café interior. White walls, gray marble-tiled floor, a tall white cylindrical pillar on the left with a large green potted plant next to it. On the wall behind her is elegant black cursive wall text that reads “let coffee connect us”. A modern black rectangular wall sconce light is mounted above the text. A dark door frame is visible on the far right. Soft, warm indoor lighting with gentle shadows.

Shot from a slightly low angle, full body visible from head to mid-calf, natural skin texture, high detail, realistic fabric sheerness and ruche folds.
Run this prompt
The pore-level beauty close-up

The pore-level beauty close-up

u/lukmanfebrianto

Skin, brows, fabric and gemstone each get a line, then light and lens. Run it at 3:4.

A tightly cropped beauty photography portrait, framed from the forehead to the chin, filling most of the frame. A young woman in her mid-20s with healthy, natural skin texture: visible fine pores across the nose, cheeks, and forehead, a soft natural sheen rather than flat smoothness, and subtle natural skin variation rather than artificial perfection. Her eyebrows are full and well-groomed, with individual hair strands visible and naturally textured. A few loose hair strands frame her hairline, each strand rendered individually and catching the light. Her eyes are deep brown, clear and bright, with fine natural texture in the iris, individually separated eyelashes, and natural moisture reflection. Her lips have a soft natural rose tone with visible fine texture and subtle natural sheen.

She wears a cream-colored linen headscarf draped loosely to one side, the fabric showing a visible tight woven pattern and soft natural folds. A small blue gemstone stud earring is visible near her ear, its facets catching the light with sharp reflections and a subtle cool blue sparkle against her warm skin tone.

Lighting: soft directional window light from the left side, creating natural shadow falloff along the cheekbone and jaw to reveal skin texture. Background is a plain, softly out-of-focus warm beige wall. Photographed with a 100mm macro lens at f/4, sharp focus on the eyes and skin detail, natural color grading, unretouched documentary-style realism, high resolution, 4:5 aspect ratio.
Run this prompt
What people say

What people said after seeing Qwen Image 3 results

Replies from public Reddit threads, quoted as written.

Qwen image 3 win

Compared

How it compares with GPT Image 2, ChatGPT Images 2.5, Qwen Image 2 and video-first

People ran these matchups themselves. Where words sit inside the image, Qwen Image 3 came out ahead.

Qwen Image 3 vs GPT Image 2

An r/ChatGPT test restyled a comic panel with three models. The poster gave Qwen the text round.

qwen was the only one who preserved the texts from the comics. None of the other ones have it.

Run the same prompt on GPT Image 2

vs ChatGPT Images 2.5

One r/generativeAI poster ran one prompt through both and picked the Qwen image.

the text is well-rendered and actually makes sense.

vs Qwen Image 2

A user who ran both on identical prompts named where 3.0 wins.

3.0 interpretation of specific details within text prompts seems to be much more accurate.

Key frames first vs text-to-video alone

The r/comfyui short drew each shot in Qwen Image 3 Pro, then animated it in MiniMax H3.

I found this much easier to control than trying to generate the whole thing directly from text.

Animate a frame in MiniMax H3
FAQ

Qwen Image 3 FAQ

The questions people ask most in Qwen Image 3 threads.

Not as of September 2026: Alibaba has published no weights, and ComfyUI's nodes call an API. Here there is nothing to install: pick the model, type your brief, generate.

You can start free: log in and Kavel adds free credits with no card, which run Kavel's own image engine. This model runs on any paid plan or credit pack, with the cost shown before you generate.

25 credits at 1K and 45 at 2K, edits included. Failed runs are refunded automatically, unless the prompt broke the content policy.

Yes. Upload the picture, write what should change, such as a new headline, and generate.

For keeping the words already on a page, one r/ChatGPT side-by-side says yes: only Qwen kept the comic's text. Both share one balance here, so test your own prompt on each.

Alibaba lists 12 languages, and people have posted Japanese and Chinese results. Quote each line and say where it sits.

Render at 2K, give each text block its exact words and position, and ask for sharp readable small text. Zoom in to judge.

Yes. Make the key frame here, then upload it to MiniMax H3 image-to-video, the route behind one 90-second r/comfyui short.

Make your first Qwen Image 3 image

Make that poster or storyboard before the idea fades. Paste the brief, pick 1K for 25 credits or 2K for 45, and leave with the image. A run that errors out on our side gives its credits back.