Logo

Image · alibaba/qwen-image-3

Qwen Image 3 — images with readable text in the frame

Qwen Image 3 is the next generation of Alibaba's image model. The Qwen-Image line is known for rendering in-frame text — signs, packaging, posters — more accurately than most competitors, and the third version continues that specialty. On limko the model works both ways: text-to-image and editing an existing picture from a reference.

from 5 ₽ per image

Where Qwen Image 3 excels

  • Frames with lettering: a shop sign, a label, a poster — where other models smear the letters.
  • Reference-based editing: upload a picture, describe the change — match_input_image keeps the aspect of the original.
  • Non-standard formats: nine aspect ratios including wide 2:1 and vertical 1:2 for banners and stories.

Qwen Image 3 or Nano Banana / GPT Image

Google's Nano Banana is the everyday workhorse for photo edits and character series. Pick Qwen Image 3 when the lettering in the frame matters: typography is the line's specialty, and it holds up better on signs, packaging and posters.

GPT Image 2 is also strong with text but noticeably pricier per run. At about one and a half credits, Qwen Image 3 is the cheapest way to get a frame with readable letters; for a final cover with dense layout, still compare both models on your prompt.

What you need as input

A text prompt; for editing — a reference image. Negative prompt (what to keep out of the frame), seed for reproducible series and prompt expansion are available — the model fleshes out a short description on its own.

Aspect ratios: 1:1, 16:9, 9:16, 4:3, 3:4, 3:2, 2:3, 2:1 and 1:2, or match_input_image to repeat the reference format. One run — one image, the price is shown on the button before you click.

Where the model is limited

One run returns one frame — for a series of variants, launch several generations with different seeds. Long phrases with complex layout are safer as short wording: that is the general rule for in-frame text with any model.

Qwen Image 3 FAQ

How does the third version differ from the previous Qwen-Image?

It is the next generation of the line: closer prompt following and more stable reference-based editing. The previous version stays in the catalog — a series started on it can continue without a style shift.

Does it understand non-English prompts?

Yes, it handles them, and it can render non-Latin lettering in the frame. For fine style wording English remains the safer choice — the general rule for catalog models.

How does reference-based editing work?

Upload the source image and describe the change. The match_input_image mode keeps the source aspect, and the negative prompt lets you explicitly forbid the unwanted.

How much does a generation cost?

About one and a half credits per frame — the exact price is always visible in the interface before launch.