Qwen Image

Advertised as qwen-image
Quick reference
Qwen Image — TLDR
  • 🏢 Alibaba Qwen team's first image generation foundation model.
  • 📏 20B-parameter Multimodal Diffusion Transformer (MMDiT) architecture.
  • 📚 Excels at complex text rendering, especially Chinese scripts.
  • 🔧 Handles both image generation and precise image editing.
  • 👁️ Also supports detection, segmentation, depth and edge estimation.
  • 🔒 Open-weights under the permissive Apache 2.0 license.
  • 🆕 Foundation for later Qwen Image 2 and Edit variants.
💰 Best price on AntSeed
$0.01067%
per generated image · cheapest seller
📏 Context
🐜 Sellers
3
advertising on AntSeed
Provider

Alibaba Group is a Chinese multinational technology company founded in 1999 and headquartered in Hangzhou, Zhejiang. Originally built around e-commerce and cloud computing, Alibaba has become one of the most prolific contributors to open-weight AI research, developing the Qwen…

Explore 39 more models by Alibaba Group
About this model

Qwen Image is the first image generation foundation model released by Alibaba's Qwen team, a 20B-parameter Multimodal Diffusion Transformer (MMDiT) open-sourced under the Apache 2.0 license. Its defining strength is high-fidelity text rendering across diverse images, preserving typographic detail, layout coherence and contextual harmony for both alphabetic languages like English and logographic scripts like Chinese. The model is more than a generator: it also supports image understanding tasks including object detection, semantic segmentation, depth and edge (Canny) estimation, novel view synthesis, and super-resolution.

As the base release, it anchors a growing family. The editing companion, Qwen Edit 2511, extends Qwen Image's text-rendering ability to instruction-based editing, feeding the input image into Qwen2.5-VL for semantic control and a VAE encoder for appearance control, enabling bilingual text edits while preserving font, size and style.

The lineage later advanced with Qwen Image 2 and its higher-fidelity Qwen Image 2 Pro tier, alongside their editing counterparts, building on the foundation that the original release established. Alibaba also exposes the family through its Model Studio text-to-image API for production use.

For users wanting the original open-weight foundation with strong compositional control and bilingual typography, Qwen Image remains the accessible Apache 2.0 starting point that the rest of the family was constructed upon.

View source on GitHub ↗View model card on HuggingFace ↗
Sources
alibabacloud.comCall Qwen Text-to-Image API via Python and Java Examples - Model Studio - Alibaba Cloud· alibabacloud.comhuggingface.coQwen/Qwen-Image-Edit · Hugging Face· huggingface.co

This About section is AI-generated from public sources via VeniceStats + Venice inference, with no human editing. It may contain inaccuracies.

Usage on AntSeed
Images generated
13
17.72k tokens · 948.00 excl. image credit
Requests
14
settled calls
Buyers
3
distinct, on this model
Sellers used
1
of 3 advertising
Settled
$0.16
gross USDC, this model
Sellers serving Qwen Image (3)compare on the network explorer →
SellerReputationRouting$ / imgCategoriesAPI

"Best price" and the seller table are live AntSeed catalog data (advertised $/1M tokens — or $ per generated image for unit-billed image models — not settled amounts). Reputation = on-chain trust (0-100). "Routing" = the SDK's default buyer routing (what the VPR desktop app ships with): a trust ≥ 60 gate on the effective reputation, then cheapest-first among routable sellers; live failover state (per-peer cooldowns) is buyer-side runtime and not included. Model knowledge (TLDR, provider, About) via the VeniceStats enrichment layer. Advertised catalog, not the model used in any specific purchase. "Usage on AntSeed" counts only settlements whose buyers share the per-model split on-chain (metadata v2/v3, opt-in), so every usage figure is a lower bound.