The Decoder→ original

Qwen-Image-3.0 от Alibaba: читаемый текст в 10 пикселей и инфографика за один проход

Команда Qwen (Alibaba) представила Qwen-Image-3.0 — генератор изображений, который разборчиво рендерит текст размером от 10 пикселей, принимает промпты до 4500 токенов и поддерживает 12 языков. За один проход модель собирает инфографику, LaTeX-статьи и газетные полосы — правда, на выходе плоская картинка, а не редактируемый макет.

AI-processed from The Decoder; edited by Hamidun News
Qwen-Image-3.0 от Alibaba: читаемый текст в 10 пикселей и инфографика за один проход
Source: The Decoder. Collage: Hamidun News.
◐ Listen to article

Qwen team (Alibaba) unveiled Qwen-Image-3.0 in July 2026 — an image generator that accepts prompts up to 4500 tokens long, renders readable text as small as 10 pixels, and natively supports 12 languages.

What Qwen-Image-3.0 can do

Qwen-Image-3.0 generates complex layouts in a single pass: infographics, pages styled like scientific LaTeX papers, and entire newspaper spreads. The main difference from previous generators is the readability of small type: the model outputs legible text as small as just 10 pixels tall, while most diffusion models turn such text into unreadable mush.

Key specs from the Qwen team:

  • Prompt length — up to 4500 tokens
  • Minimum readable text — from 10 pixels
  • Native support for 12 languages
  • In a single pass — infographics, LaTeX-style papers, newspaper spreads

The catch with complex layouts

The practical value of such layouts remains questionable. Qwen-Image-3.0 delivers a finished picture, not an editable document: a generated infographic or newspaper spread can't be opened and fixed — only redrawn with a new prompt. For real layout work, where text gets edited dozens of times, this limitation matters more than the mere fact that 10-pixel type is readable.

"The practical value of such layouts is questionable when the output is a pixel image rather than an editable format," notes

The Decoder.

That said, support for 4500-token prompts genuinely expands control: a single request can describe page structure, block text, and style all at once, without splitting the task into several steps.

How this looks against the market

Qwen-Image-3.0 continues Alibaba's line of open models, which compete with closed generators like Google Nano Banana and OpenAI's solutions. The bet on in-image text and multilingual support across 12 languages is a direct answer to the weak spot of most image models: until now, text on generated images had to be fixed by hand or pasted over it in an editor.

What this means

Image generators are edging into a niche once reserved for designers and layout artists: posters, infographics, covers dense with text. As long as the output stays a "flat" picture without layers, Qwen-Image-3.0 is more a tool for drafts and previews than a replacement for an editable layout — but it raises the bar for small-text readability noticeably.

Frequently asked questions

What can Qwen-Image-3.0 do out of the box?

In a single pass, the model generates infographics, pages styled like LaTeX papers, and newspaper spreads, accepts prompts up to 4500 tokens, and renders readable text from 10 pixels in 12 languages.

Why is complex layout generation criticized?

Because the result is a pixel image, not an editable file. A finished infographic or newspaper page can't be edited like a document — only regenerated entirely with a new prompt.

ZK
Hamidun News
AI news without noise. Daily editorial selection from 50+ sources. A product by Zhemal Khamidun, Head of AI at Alpina Digital.

Need AI working inside your business — not just in your newsfeed?

I build production AI for companies — custom CRM, internal tools, autonomous agents, workflow automation. Owned by you, shaped to your process, no per-seat tax. Built by Zhemal Khamidun, CPO of AlpinaGPT (AI platform, 6,000+ users).

What do you think?
Loading comments…