Model comparison
Qwen-Image vs Gemini 2.5 Pro
Capability radar, value bars, and side-by-side posture.
Qwen-Image
Not scored
value score
Gemini 2.5 Pro
93
value score
Qwen-Image ctx
—
Gemini 2.5 Pro ctx
1.0M
Capability radar
Value · context · multimodal · openness · speed posture
Value head-to-head
- Gemini 2.5 Pro93
Attribute tape
| Field | Qwen-Image | Gemini 2.5 Pro |
|---|---|---|
| Developer | Qwen | Google DeepMind |
| Context | — | 1.0M |
| Modalities | Text, Code, Image, Audio, Video | |
| Openness | — | Proprietary Model |
| Speed | — | — |
| Price | — | $1.25/1M in · $10/1M out |
| Value score | — | 93 |
| Summary | Qwen-Image is a generative AI foundation model from Alibaba Cloud's Qwen team that creates and edits images from text prompts, with particular strength in rendering readable multilingual text (including Chinese and English) inside generated images. It also supports broader visual tasks such as style transfer, object add/remove, depth estimation, and novel-view synthesis. | Google frontier multimodal model. |