Rich content: Supports input of up to 4.5k tokens and dense information layout with images-within-images, enabling complex layouts like newspapers, storyboards, menus, and exam papers to be generated in a single pass. Authentic detail: Supports precise rendering of text as small as 10px, and vividly reproduces fine details such as micro-expressions, pores, and individual strands of hair—approaching the quality of real photography. Deep knowledge: Supports native rendering of 12 languages and 20+ fonts, realistic simulation of mainstream interfaces such as web pages, games, and live streams, fully incorporating external knowledge. Qwen-Image-3.0-Pro isn't just pursuing "good looks"—it's pursuing "usefulness", making image generation a truly deployable productivity tool.
Model Capabilities
| Capability | Support | Capability | Support |
|---|---|---|---|
Input Modality |
Image Text |
Output Modality |
Image |
Model Experience |
Function Calling |
||
Structured Outputs |
Web Search |
||
Prefix Completion |
Context Caching |
||
Batch Inference |
Fine-tuning |
Context Limits
| Parameter | Value | Parameter | Value |
|---|---|---|---|
Max Input Length |
— |
Max Output Length |
— |
Context Window |
— |
Pricing
No public pricing information available.
Rate Limits
China (Beijing)
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) |
1 |
Singapore
Scope: International
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) |
1 |