Qwen Image 3 AI Image Generator

Qwen Image 3 is Alibaba's working-tool model: menus, posters and documents where the text inside the picture has to be spelled right.

0 / 2,000
~20 cr

What is Qwen Image 3?

Qwen Image 3 is the standard tier of Alibaba's 2026 image family, released on 21 July 2026 and API-only — unlike the earlier Qwen-Image releases, there are no published weights and no technical report. Alibaba pitches it as a working tool, not an art toy, and the strength reviewers keep landing on is in-image text: Eden AI scored its text rendering 8.5/10, and our five text tests below all came back correctly spelled on the first take.

It is not the model for everything. Our chart test reproduced the failure independent testers found: the numbers print correctly, but the axis is out of order, the bars are not proportional to their values, and it invented a citation for data we made up. Reviewers have also documented Korean misspellings — so the multilingual claim is really a CJK-and-English claim — and occasional anatomy errors in figures.

Release details and positioning are from independent coverage of the launch: Unite.AI — Alibaba launches Qwen Image 3 without benchmarks or weights

What reviewers say

The model can write readable text inside generated images… scores 8.5/10 on text rendering, outperforming DALL-E 3 (7.0) and SD 3.5 (6.5).
Eden AI — model comparison
The biggest improvement is that it understands long and detailed instructions while producing highly accurate images with readable text.
Mehul Gupta — Data Science in Your Pocket
A subpar equivalent to other proprietary models like GPT Image 2 and Nano Banana Pro.
BuildFastWithAI — hands-on review

Every image below was generated on this page on 19 August 2026, with the exact prompt shown — including the chart test it failed. Nothing was retouched or picked from a batch.

Menus, posters and price lists that come back spelled right

This is the model's home ground. We gave it a printed diner menu with three sections and seven prices at small type sizes, an event poster with a headline, a schedule line and a small footer, and a bookstore poster with a vertical Chinese title over an English subtitle. All of it came back correctly spelled and correctly set — including things we never asked for: the menu grew a plausible "Est. 1952" and the poster a whole sidebar of extra labels, all spelled right, because automatic prompt rewriting is on by default and enriches short briefs. If you want only your text and nothing else, say so in the prompt.

Menu — every price correct
Qwen Image 3 diner menu for The Bluebird Diner with breakfast, lunch and drinks sections, all prices legible and correct2K · 2:3 · 113s
English — correct
Qwen Image 3 Harvest Market event poster with heading, schedule line and footer all correctly spelled2K · 3:4 · 105s
Chinese — correct
Qwen Image 3 Chinese bookstore poster with the vertical title 夜航书局 and English subtitle rendered correctly2K · 2:3 · 95s

The test is not the headline — headlines are easy. It is the 7.50 and the 6.80 holding at menu-item size, and both scripts landing in one workspace. One caveat from the same batch: the model happily adds correctly-spelled content you did not ask for, so brief it tightly when the layout is fixed.

Attach a photo, keep the label, change the world around it

The same model edits. We generated a product shot of a labelled tea caddy, handed it back as a reference, and asked for the identical tin at an outdoor winter market at dusk. The tin, its shape, its colour and both lines of label text survived the move; the lighting was rebuilt for the new scene, down to a warm highlight on the metal against cold blue snow light. Up to 3 reference images per run — enough for product-in-scene work, not for the many-reference composites the Nano Banana channels are built for.

Prompt

Studio product photograph of a matte sage-green tea caddy on pale travertine, its label reading 'KIYOMI' in a fine serif with a smaller line 'ceremonial grade matcha · 40g' beneath, soft directional daylight from the right, one clean specular edge highlight, shallow depth of field, e-commerce hero framing.

Qwen Image 3 studio product photograph of a sage-green tea caddy with the label KIYOMI ceremonial grade matcha 40g
2K · 1:1 · 101s

Keep the same sage-green tea caddy from the reference image exactly as it is — same label text 'KIYOMI', same shape and colour — but place it on a weathered wooden stall at an outdoor winter market at dusk, light snow falling, warm string lights bokeh behind, cold blue ambient light with one warm highlight on the tin.

Qwen Image 3 edited output: the identical tea caddy on a market stall at dusk with snow falling, label still reading KIYOMI
2K · 1:1 · 79s · 1 reference image

One honest nit from step 1: the label text is correct but its line hierarchy shifted — "ceremonial" moved up beside the brand name instead of leading the second line. The edit then preserved that layout faithfully, wrong line break and all.

The pricing quirk: 2K costs exactly what 1K costs

This is where the standard tier genuinely differs from its Pro sibling. On Qwen Image 3 Pro, 2K costs more than 1K; here they are priced identically, so this page defaults to 2K and the only thing 1K buys you is a faster loop. We ran the same typewriter macro at both tiers: 57 seconds at 1K, 84 at 2K, with roughly four times the pixels at 2K. Iterate at 1K if you are impatient; switch to 2K for anything you keep.

1K
Vintage typewriter keys macro generated at 1K in 57 seconds1216×832 · 57s
2K
The same typewriter keys prompt generated at 2K in 84 seconds2496×1664 · 84s

Same prompt, same settings, only the resolution tier changed. The cost is the same either way — which is the whole point of this model, and the reason 2K is the default here.

The chart test — where we stop trusting it

Independent testers reported that its data charts look right and are wrong, so we ran our own: a five-bar chart with values we supplied. Every label and every number printed correctly — and the geometry underneath them is broken. The axis repeats one value and puts another out of order, the tallest bar tops out below where its printed value says it should, and the footer credits a "Global Coffee Report 2025" that does not exist, for data we invented. Treat it as a typographer, not an analyst: it will set your chart title beautifully, but build real charts in a charting tool.

Failed — axis and proportions wrong
Qwen Image 3 coffee harvest infographic: values 3.7, 1.8, 0.8, 0.7 and 0.5 all printed correctly, but the y-axis repeats 1.5, misorders 2.0 and 2.5, and the bar heights do not match the numbers2K · 3:4 · 118s

Look closely at the left axis: 1.5 appears twice and 2.0 sits above 2.5. The prompt asked for bars "strictly proportional" to their numbers; they are not. A page that only printed the successes would not tell you this.

The rest of what we ran

The groups above test specific claims; these four are the breadth check from the same 19 August 2026 batch, each first take with the exact prompt. Across all 14 completed runs, 2K generations took 64–118 seconds with a median of 84. One submission failed at the upstream provider before generating; re-running it unchanged succeeded in 75 seconds — that retried image is the breakfast tray below.

Qwen Image 3 ultra-wide 21:9 panorama of a harbor town at first light with fishing boats and mist
Prompt2K · 21:9 · 3072×1312 · 93s — it added a small harbor sign we never asked for, correctly spelled

Ultra-wide cinematic frame: a small harbor town at first light, fishing boats resting on glassy water in the lower third, strings of unlit bulbs sagging between mooring posts, mist rolling off the headland behind, soft peach-and-slate palette, anamorphic depth, fine detail across the full width.

Qwen Image 3 documentary portrait of an elderly fisherman mending a green net at dawn with detailed skin texture
Prompt2K · 2:3 · 75s — hands and face came back correct in this run, though reviewers do report anatomy slips

Documentary portrait of an elderly fisherman mending a green net at dawn, weathered hands and deeply lined face clearly rendered, salt-and-pepper stubble, warm low sun from the left against a cold overcast sea, realistic skin texture without smoothing, shallow depth of field.

Qwen Image 3 overhead photograph of a Chinese breakfast tray with congee, youtiao, pickles and soy milk
Prompt2K · 3:2 · 75s — the retried run; the first submission failed upstream without generating

Overhead editorial food photograph of a lacquered tray breakfast: a bowl of congee with scallions and white pepper, a plate of golden youtiao, pickled radish in a small dish and a glass of soy milk, steam rising, warm morning window light, rich texture in the ceramics and wood.

Qwen Image 3 isometric illustration of a night market street with glowing stalls and lantern strings
Prompt2K · 1:1 · 75s

Isometric illustration of a small night market street: glowing food stalls with striped canopies, tiny figures queueing, lanterns strung overhead, steam and smoke drifting up, warm oranges against deep blue night, clean vector-adjacent rendering with soft gradients, no text on any sign.

Who Qwen Image 3 is for

The standard tier is the volume workhorse of the Qwen pair. Each of these maps to something demonstrated further up the page.

  • Cafés, restaurants and local businesses

    Menus, price boards and event flyers where every price has to be right on the first take. Our seven-price menu test is exactly this job, and it passed.

  • Marketers working in English and Chinese

    One workspace for both scripts — our bookstore poster set a vertical Chinese title and an English subtitle in the same frame, both correct. Proofread any script you cannot read yourself; reviewers caught Korean errors.

  • Anyone iterating in volume

    A flat price whatever the resolution, so exploring ten directions costs what it costs — no tier arithmetic. Draft at 1K for speed if you like, keep the finals at 2K for free.

  • Product teams reusing existing shots

    Attach up to 3 references and move a product into a new scene with its label text intact, as in our winter-market edit. For composites with many references, use a Nano Banana channel instead.

  • Not for: charts carrying real data

    Our chart test printed the right numbers on broken geometry and invented a source. Set real data in a charting tool, then let this model do the poster around it.

Qwen Image 3 vs Qwen Image 3 Pro vs GPT Image 2 vs Z Image Turbo

The three models you would realistically weigh against this one — same workspace, same subscription, one click to switch.

Qwen Image 3Qwen Image 3 ProGPT Image 2Z Image Turbo
Credits per image20, flat — 1K and 2K identical30 at 1K · 50 at 2K5 – 560, depending on settings0 — free
Cost changes with resolutionNoYes — 2K costs moreYes — 1K/2K/4K × low/medium/highNo resolution tiers
Max resolution2K2K4K1K-class presets
Reference imagesUp to 3Up to 3Up to 16None — text to image only
Aspect ratios8, up to 21:98, up to 21:98 plus auto, up to 21:99, up to 2:1 and 1:2
Images per run11Up to 101
In-image text, our own testsMenu, posters and bilingual title all correct; charts unreliableNot measuredThe most reliable in our tests on its pageNot tested
Our measured speed64–118s at 2K, median 84s (13 runs)Not measured32s low · 68s medium · 166s high at 2K (its page, single runs)Not measured

Credits, reference-image limits, resolutions and ratio counts are read from our own model registry on 19 August 2026 — the GPT Image 2 ratio count excludes "auto", which is a setting rather than a ratio. Speeds are our own measurements with the sample size stated; cells we have not tested say so rather than borrowing a vendor number.

Explore More Models

Compare Qwen Image 3 with other image models — same workspace, one click to switch.

Qwen Image 3 specs on CreateVision AI

Modelqwen3/text-to-image and qwen3/image-to-image, by Alibaba's Qwen team (released July 21, 2026; API-only, no open weights)
ModesText-to-image and image-to-image, including reference-guided editing
Quality tiers1K and 2K — priced identically, so 2K is the default here
Aspect ratios8 presets from 1:1 to 21:9 ultra-wide
Images per runOne
Reference imagesUp to 3 per generation
Prompt controlsNegative prompt, seed, and automatic prompt rewriting (on by default) — prompts up to 5,000 characters
Text renderingEnglish and Chinese correct across our five text tests; reviewers report Korean errors; chart-style data layouts unreliable (see our failed test above)
Measured speed64–118s at 2K, median 84s (13 runs); 57s on our single 1K run — 19 August 2026
Reliability13 of 14 first submissions completed; one failed at the upstream provider and succeeded when re-run unchanged

How to use Qwen Image 3 online

Start Generating Images
  1. Open the workspace

    Create a free account — no API key, no setup. The generator on this page is already set to Qwen Image 3.

  2. Put your text in quotes

    Write the exact words you want rendered — headings, prices, a Chinese title — in quotation marks. If the layout is fixed, say that only this text should appear: prompt rewriting is on by default and will otherwise invent extra, correctly-spelled content.

  3. Attach up to 3 references to edit

    Hand it a photo and describe the change. Label text on the reference survives the edit; the lighting is rebuilt for the new scene.

  4. Pick a ratio and leave 2K on

    Eight ratios up to 21:9. 1K and 2K cost the same, so 2K is the default — the exact credit cost shows before every run.

Related reading

Choose Your Plan

Start free with welcome credits. Upgrade for premium models, faster generation, and 4K output.

Free

$0

Perfect for getting started

  • 200 welcome credits, up to 80/day
  • Z Image Turbo drafts at 0 credits
  • Standard generation speed
  • Generation history included
Get Started

Premium

Most popular
$29/month

Billed annually · $240/year

  • 8,000 credits/month, no daily limits
  • All premium models — GPT Image 2, Nano Banana, Seedream, Seedance, Veo, Kling
  • 5x generation speed, priority queue
  • AI prompt enhancement
Upgrade to Premium

Ultimate

$49/month

Billed annually · $380/year

  • 18,000 credits/month, no daily limits
  • HD generation up to 4K resolution
  • Priority support
  • Fastest queue and early access to new models
Upgrade to Ultimate

Frequently Asked Questions

1

Is Qwen Image 3 free to use?

The model itself is paid, at a low flat rate per image that is identical at 1K and 2K — the exact credit cost is shown before every run. Every account gets free daily credits, and Z Image Turbo (0 credits) is there for throwaway drafts before you spend anything.

2

Qwen Image 3 vs Qwen Image 3 Pro — which should I pick?

Start here. The standard tier already handles menus, posters, bilingual titles and reference edits, and its flat price makes volume work predictable. Move to Pro when the job is fine detail at maximum polish — it charges separately for 2K, so it makes sense for the frames you ship, not the ones you explore.

3

Is Qwen Image 3 better than GPT Image 2?

For most everyday text work, it gets you there at a flat, predictable price — our menu, poster and bilingual tests all passed first take. GPT Image 2 remains the more dependable choice when dense small type or a complex multi-element brief must be right in a single frame, at a cost that varies from 5 to 560 credits by settings. Volume goes here; the one frame that cannot be wrong goes there.

4

Can it render Chinese text — and other languages?

Chinese and English are its strong suit: our bookstore poster set a vertical Chinese title with an English subtitle, both correct, and typeset the way a local designer would. Beyond CJK and English, be careful — independent testers documented Korean output with mixed-up vowels and misspelled words. Proofread any script you cannot read, or set that type yourself.

5

How long does a generation take?

In our 19 August 2026 batch, 2K runs took 64 to 118 seconds with a median of 84; our single 1K run took 57. Generation is asynchronous — the result lands in your history with its prompt and parameters saved, so a good setup is reproducible later.

6

What is Qwen Image 3 bad at?

Three things from our own runs and the published record. Charts with real data: ours printed correct values on a broken axis with non-proportional bars, and invented a source. Unrequested additions: prompt rewriting is on by default and will enrich a short brief with extra (correctly spelled) content — constrain it in the prompt for fixed layouts. And figures: our portrait run was clean, but reviewers report extra limbs and inconsistent eyes often enough to warrant a check before you ship a person.

7

Can I use Qwen Image 3 images commercially?

Yes, images you generate can be used in personal and commercial projects under our terms of service. You remain responsible for prompt content, trademarks, and the rights to any reference images you attach.

Put a menu, a poster or a bilingual title in front of Qwen Image 3

Start Generating Images

Discover our one-click AI image tools

One-click tools for everything around generation — polish, restyle, and share.

Pricing & Access — The FactsQwen Image 3

Model
Qwen Image 3
Developer
Alibaba
Access
Runs in the CreateVision AI workspace via API — sign in and generate. No API key, no cloud project, no software install.
Price per generation
From 20 credits ≈ $0.07 per image on the Premium plan ($29/month for 8,000 credits). The free plan includes 80 credits every day.
Privacy
Private by default — no other user can see your generations, on any plan. We don't sell or misuse member data. Privacy Policy