Qwen Image 3 Pro AI Image Generator

Qwen Image 3 Pro is the flagship of Alibaba's Qwen-Image line, and the reason to run it is bilingual type — Chinese and English in one frame.

0 / 2,000
~30 cr

What is Qwen Image 3 Pro?

Qwen Image 3 Pro is the flagship tier of Alibaba's Qwen-Image 3.0 line, released on 21 July 2026. Alibaba aims it at information-dense work — long briefs, layouts, and in-image text at very small sizes — but shipped it with no technical report, no model card and no downloadable weights, so arena scores and independent tests are the only third-party evidence there is.

It is not the best tool for everything. Independent testers report Korean text with mixed-up vowels, charts that look polished while getting data wrong, and portraits that trail the photoreal specialists. Our own ten runs mostly landed — but the small-print test below collapsed mid-paragraph, and two of the words that stayed legible came out misspelled. For maximum skin and material realism use Seedream 5 Pro; for the most dependable small type, GPT Image 2.

Arena standing is from the Artificial Analysis text-to-image leaderboard, checked 19 August 2026: Artificial Analysis leaderboard

What reviewers say

The most capable text-in-image model Alibaba has built.
Build Fast with AI — Qwen-Image 3.0 review
Legible text as small as 10px, photographic skin and hair texture.
Yash Thakker, explainx — on the launch claims
A subpar equivalent to other proprietary models like GPT Image 2 and Nano Banana Pro.
Yash Thakker, explainx — community roundup

Every image below was generated on this page on 19 August 2026 — ten prompts, ten first-take results, 57 to 93 seconds each, with the exact prompt shown. Including the test that exposed a real limit.

Qwen Image 3 Pro concert poster with the vertical Chinese title 松风入弦 rendered correctly beside a guqin under a spotlight

Chinese typography it sets like a native

This is the home ground of the Qwen image family. We asked for a concert poster with a four-character Chinese title running vertically down the left in a songti face, and got exactly that: every character correct, the stroke weight consistent, the English subtitle spelled right in spaced capitals — and a red seal chop by the title that was not in the brief, which is precisely what a designer trained on Chinese poster convention would add. The bilingual tea tin further down this page came back with both scripts correct too. If your work switches between English and Chinese in one frame, this is the strongest reason to pick this model.

Qwen Image 3 Pro diner menu board with three sections of correctly spelled dishes and prices

Three type sizes on one board, every word correct

A menu board is a harder text test than a headline: three sections, six dish names, six prices and a small footer all have to hold in one frame. Ours came back with every word and every price correct, from the hand-painted heading down to 'Served till 11am'. One quirk worth knowing: our brief separated dishes with slashes, and the model painted those slashes onto the board as literal marks. It follows text instructions almost too literally — do not put punctuation in quotes unless you want it rendered.

Qwen Image 3 Pro recipe card where the middle of the small-print paragraph degrades into unreadable marks

Where the small-print claim breaks: a full paragraph

Alibaba's launch claims include legible text at 10 pixels. So we asked for a vintage recipe card carrying a full forty-word method paragraph in small type. The result is the most honest image on this page: the title and the first five lines are real, readable English, then the middle of the paragraph collapses into ink-shaped marks that only look like words, before recovering cleanly for the final seven lines. Two of the legible words are also misspelled — 'soffened' and 'toothpack'. Display-size text is reliable here; long small print is not. Keep paragraphs off the image, or set them yourself afterwards.

The same brief at 1K and 2K

The Pro tier prices 1K and 2K separately, so we ran the identical prompt once at each. Both frames keep genuinely weathered skin — creases, stubble, sun damage — rather than the smoothed faces this model family gets criticised for, and the 2K frame resolves visibly finer net and rope detail. Both runs also did something we never asked for: they wrote signage into the scene, 'PIER 7' on a mooring post in one and 'HARBOR FISHERIES' on a crate in the other, correct at display size while the smaller line beneath the crate stencil degraded. That is this model's character in one pair: it volunteers text constantly, and the size threshold decides whether the text survives.

1K
Qwen Image 3 Pro fisherman portrait at the 1K tier, with a self-added PIER 7 sign rendered correctly1K · 3:2 · 76 seconds
2K
Qwen Image 3 Pro fisherman portrait at the 2K tier, with a self-added HARBOR FISHERIES crate rendered correctly2K · 3:2 · 93 seconds

Same prompt, both first take. We did not fix the seed, so the compositions differ — what is comparable is the rendering, not the framing. For a like-for-like variation workflow, fix the seed in the advanced controls.

Qwen Image 3 Pro bar chart rendering five specified values correctly, with a self-added correct total

We probed the reported chart weakness — this run passed

The sharpest criticism in independent reviews is data integrity: charts that look right while being numerically wrong. We probed it directly — a bar chart brief with five exact values and strict proportionality. The result printed all five numbers correctly, kept Thursday tallest and Friday shortest, and then went beyond the brief: it added a 'Total Cups Sold: 950' line, which is the correct sum of our five values, and peak-day callouts that match the data. The one artifact is a legend entry for a 'Daily Target' series that appears nowhere in the chart. One passed probe does not retire the criticism — the reviewers' failures are documented — but it does mean the failure is not guaranteed. Check every number before a chart ships; on this image that check takes a minute.

The rest of what we ran

Four more first-take outputs from the same 19 August 2026 batch, each with the exact prompt — macro detail, bilingual packaging, food and a 21:9 ultra-wide. Across all ten runs: 57 seconds at the fastest, 93 at the slowest, median 79.

Qwen Image 3 Pro macro photograph of a mechanical watch movement with a self-added engraving rendered correctly
Prompt2K · 3:2 · 90 seconds · the engraved lettering on the bridge was the model's own addition, correctly spelled

Extreme macro photograph of an open mechanical watch movement: polished gears, blued screws, a coiled balance spring and engraved bridges, one warm raking light from the left revealing machining marks on the metal, shallow depth of field, jewel bearings glinting ruby red.

Qwen Image 3 Pro tea tin with the Chinese label 云雾龙井 and its English subtitle both rendered correctly
Prompt2K · 1:1 · 69 seconds · both scripts correct

Studio product photograph of a tea tin on pale linen: the cylindrical tin's label reads '云雾龙井' in refined Chinese calligraphy across the centre, with 'MISTY DRAGONWELL — GREEN TEA' in small spaced capitals beneath it and a tiny line 'NET WT 100g' near the base, sage green and cream label design, soft window light, both scripts sharp and correctly written.

Qwen Image 3 Pro overhead food photograph of hand-pulled noodles, with a self-added menu card spelled correctly
Prompt2K · 3:2 · 80 seconds · the menu card is another unprompted addition, every word correct

Overhead editorial food photograph of a lacquer tray holding a bowl of hand-pulled noodles in clear broth, chopsticks resting on a ceramic stand, small dishes of chilli oil and pickled greens beside it, steam rising, warm side light on a dark slate table, crisp texture in the noodles and scallions.

Qwen Image 3 Pro 21:9 ultra-wide night scene of a sleeper train crossing an illuminated bridge over a river
Prompt2K · 21:9 · 93 seconds · it named the train and lit a skyline logo on its own — legible at display size, degrading on the smallest carriage labels

Ultra-wide cinematic frame: a night train crossing an illuminated bridge over a calm river, city lights doubled in the water, cool blue night tones against the warm carriage windows, anamorphic composition with the bridge running the full width of the frame.

Who Qwen Image 3 Pro is for

Each of these maps to something demonstrated on this page — not a wish list.

  • Bilingual brands and cross-border sellers

    Packaging, posters and store assets that carry Chinese and English in the same frame, both set correctly — the tea tin and guqin poster above are exactly this job. Have a native reader proofread anything customer-facing.

  • Menu boards, signage and price lists

    Three type sizes, six prices and a footer line all held in our menu test. Keep the smallest text at label size, not paragraph size, and it stays readable.

  • Detail-heavy product and editorial shots

    The watch-movement macro and the noodle tray show what the Pro tier buys: coherent small structure — gear teeth, steam, scallions — instead of smoothed-over texture.

  • Infographic and layout drafts

    Our five-value chart probe came back numerically correct, including a self-added total. Treat every number as unverified until you have read it — the failure mode reviewers describe is real, it just is not guaranteed.

  • Not for: paragraphs on the image, or scripts you cannot read

    Our forty-word paragraph collapsed mid-way, and independent testers report vowel-level errors in Korean. Set long copy yourself, and proofread any script you cannot verify by eye.

Qwen Image 3 Pro vs Qwen Image 3 vs GPT Image 2 vs Seedream 5 Pro

Same workspace, same subscription — switching between any of these is one click.

Qwen Image 3 ProQwen Image 3GPT Image 2Seedream 5 Pro
Credits per image30 at 1K · 50 at 2K20 flat — 1K and 2K cost the same5 to 560, by resolution × quality (+10 per reference image)50 at 1K · 100 at 2K
Max resolution2K2K4K2K
Reference images33165
Images per run11Up to 101
Aspect ratios8, from 1:1 to 21:98, from 1:1 to 21:99, including auto8 plus auto, or exact custom pixels
In-image text, our own testsDisplay text correct in five tests; a small-print paragraph collapsedNot testedThe most reliable we have measuredDisplay type strong; CJK correct, Arabic misspelled
Our measured speed57–93s, median 79s (10 runs)Not measured32s low · 68s medium · 166s high at 2K (its page)~87s at 1K, ~107s at 2K (15 runs)

Credits, reference-image limits, resolutions and ratio counts are read from our own model registry on 19 August 2026 — GPT Image 2 prices on a resolution × quality matrix, which is why its cell is a range rather than one number. Speeds are our own measurements with the sample size stated, taken on each model's own page. Cells we have not tested say so rather than borrowing a vendor number.

Explore More Models

Compare Qwen Image 3 Pro with other image models — same workspace, one click to switch.

Qwen Image 3 Pro specs on CreateVision AI

Modelqwen3/pro-text-to-image and qwen3/pro-image-to-image, by Alibaba's Qwen team (Qwen-Image 3.0 line, released July 21, 2026)
Arena standing8th on the Artificial Analysis text-to-image arena, Elo 1284 — checked August 19, 2026
ModesText-to-image, and reference-guided editing with up to 3 images
Quality tiers1K and 2K
Aspect ratios8 presets, from 1:1 to 21:9 ultra-wide
Images per runOne
Prompt controlsNegative prompt, fixed seed, and automatic prompt expansion (on by default); prompts up to 5,000 characters
Text renderingChinese and English display text correct in all five of our text tests; a forty-word small-print paragraph partially collapsed
Credits30 per 1K image · 50 per 2K image
Measured speed57–93 seconds, median 79 — our 10 generations on August 19, 2026

How to use Qwen Image 3 Pro online

Start Generating Images
  1. Open the workspace

    Create a free account — no API key, no Alibaba Cloud console. The generator on this page is already set to Qwen Image 3 Pro.

  2. Put the exact text in quotes

    Write the scene, then quote any words that must appear on the image, in English or Chinese. It renders quoted text almost too literally — including punctuation — so quote only what you want painted.

  3. Pick ratio and tier

    Eight aspect ratios up to 21:9, then 1K to iterate or 2K for the final. The exact credit cost shows before you run.

  4. Iterate with seed and negative prompt

    Fix the seed to rerun controlled variations of a result you like, and state what to avoid in the negative prompt. History keeps every prompt and parameter.

Related reading

Choose Your Plan

Start free with welcome credits. Upgrade for premium models, faster generation, and 4K output.

Free

$0

Perfect for getting started

  • 200 welcome credits, up to 80/day
  • Z Image Turbo drafts at 0 credits
  • Standard generation speed
  • Generation history included
Get Started

Premium

Most popular
$29/month

Billed annually · $240/year

  • 8,000 credits/month, no daily limits
  • All premium models — GPT Image 2, Nano Banana, Seedream, Seedance, Veo, Kling
  • 5x generation speed, priority queue
  • AI prompt enhancement
Upgrade to Premium

Ultimate

$49/month

Billed annually · $380/year

  • 18,000 credits/month, no daily limits
  • HD generation up to 4K resolution
  • Priority support
  • Fastest queue and early access to new models
Upgrade to Ultimate

Frequently Asked Questions

1

What is Qwen Image 3 Pro best at?

In-image text at display size and information-dense layouts, in English and Chinese together. It is the flagship of Alibaba's Qwen-Image line and sits eighth on the Artificial Analysis text-to-image arena at Elo 1284. All five of our display-text tests — poster, menu, packaging, plus two pieces of signage it added on its own — came back correctly spelled.

2

Qwen Image 3 Pro vs Qwen Image 3 — which should I pick?

Draft on the standard model: it costs a flat 20 credits and renders 2K at no extra charge. Move to Pro when the result needs finer texture or has text in it that matters. The controls are identical, so the same prompt and references carry over unchanged.

3

Is Qwen Image 3 Pro better than GPT Image 2?

For long small print, no — our forty-word paragraph test collapsed here, and GPT Image 2 remains the most dependable small-type model we have measured; it also goes to 4K and takes 16 reference images. For Chinese or bilingual typography set the way a local designer would set it, Qwen Image 3 Pro is the stronger pick, at a flat, predictable price per tier.

4

How long does a generation take?

Across our ten runs on 19 August 2026: 57 to 93 seconds, median 79. Generation is asynchronous — you can keep working, and the result lands in history with its prompt and parameters saved.

5

What is Qwen Image 3 Pro bad at?

Three things worth planning around. Long small print: our forty-word paragraph came back half gibberish, with two legible words misspelled. Scripts you cannot proofread: independent testers documented vowel-level errors in Korean. Data graphics: reviewers report charts that look right with wrong numbers — our own five-value probe passed, but verify every figure before shipping. It also adds unrequested text to scenes constantly; usually correct at display size, but check for it.

6

Does it support negative prompts, seeds and prompt rewriting?

All three. State what to avoid in the negative prompt, fix the seed for controlled variations, and leave the automatic prompt expansion on to have short briefs enriched — or switch it off when you want your exact wording followed.

7

How much does Qwen Image 3 Pro cost?

30 credits per image at 1K and 50 at 2K, one image per run. The exact cost for your selected tier is always shown on the generate button before you spend anything.

8

Can I use the images commercially?

Yes, images you generate can be used in personal and commercial projects under our terms of service. You remain responsible for prompt content, trademarks, and likeness rights.

Put a bilingual brief in front of Qwen Image 3 Pro

Start Generating Images

Discover our one-click AI image tools

One-click tools for everything around generation — polish, restyle, and share.

Pricing & Access — The FactsQwen Image 3 Pro

Model
Qwen Image 3 Pro
Developer
Alibaba
Access
Runs in the CreateVision AI workspace via API — sign in and generate. No API key, no cloud project, no software install.
Price per generation
From 30 credits ≈ $0.11 per image on the Premium plan ($29/month for 8,000 credits). The free plan includes 80 credits every day.
Privacy
Private by default — no other user can see your generations, on any plan. We don't sell or misuse member data. Privacy Policy