Alibaba unveils Qwen-Image, a 20B image model with advanced text rendering

Published 04/08/2025, 17:10
© Reuters

Investing.com -- Alibaba (NYSE:BABA) has released Qwen-Image, a 20B MMDiT image foundation model that delivers significant advances in complex text rendering and precise image editing capabilities.

The new model, which users can access through Qwen Chat by selecting "Image Generation," features superior text rendering abilities that handle multi-line layouts, paragraph-level semantics, and fine-grained details. It supports both alphabetic languages like English and logographic languages such as Chinese with high fidelity.

Qwen-Image also offers consistent image editing through an enhanced multi-task training paradigm, achieving exceptional performance in preserving both semantic meaning and visual realism during editing operations.

According to Alibaba, the model outperforms existing solutions across multiple public benchmarks for both generation and editing tasks, including GenEval, DPG, OneIG-Bench, GEdit, ImgEdit, and GSO. It particularly excels in text rendering benchmarks such as LongText-Bench, ChineseWord, and TextCraft, where it significantly outperforms current state-of-the-art models.

The company demonstrated Qwen-Image’s capabilities through various examples, showcasing its ability to render complex text in different scenarios. These include accurately generating Chinese characters on shop signs with proper depth of field, creating detailed English text on book covers and information slides, and handling bilingual content with ease.

Beyond text processing, Qwen-Image supports a wide range of artistic styles from photorealistic scenes to impressionistic paintings, and offers various editing operations including style transfer, additions, deletions, detail enhancement, text editing, and character pose adjustment.

Alibaba stated that Qwen-Image aims to promote the development of image generation, lower technical barriers to visual content creation, and inspire innovative applications. The company is inviting community participation and feedback to build "an open, transparent, and sustainable generative AI ecosystem."

The model is scheduled for launch in August 2025.

This article was generated with the support of AI and reviewed by an editor. For more information see our T&C.

Latest comments

Risk Disclosure: Trading in financial instruments and/or cryptocurrencies involves high risks including the risk of losing some, or all, of your investment amount, and may not be suitable for all investors. Prices of cryptocurrencies are extremely volatile and may be affected by external factors such as financial, regulatory or political events. Trading on margin increases the financial risks.
Before deciding to trade in financial instrument or cryptocurrencies you should be fully informed of the risks and costs associated with trading the financial markets, carefully consider your investment objectives, level of experience, and risk appetite, and seek professional advice where needed.
Fusion Media would like to remind you that the data contained in this website is not necessarily real-time nor accurate. The data and prices on the website are not necessarily provided by any market or exchange, but may be provided by market makers, and so prices may not be accurate and may differ from the actual price at any given market, meaning prices are indicative and not appropriate for trading purposes. Fusion Media and any provider of the data contained in this website will not accept liability for any loss or damage as a result of your trading, or your reliance on the information contained within this website.
It is prohibited to use, store, reproduce, display, modify, transmit or distribute the data contained in this website without the explicit prior written permission of Fusion Media and/or the data provider. All intellectual property rights are reserved by the providers and/or the exchange providing the data contained in this website.
Fusion Media may be compensated by the advertisers that appear on the website, based on your interaction with the advertisements or advertisers
© 2007-2025 - Fusion Media Limited. All Rights Reserved.