Model comparison
Stable Diffusion (CompVis) vs Gemini 2.5 Pro
Capability radar, value bars, and side-by-side posture.
Stable Diffusion (CompVis)
Not scored
value score
Gemini 2.5 Pro
93
value score
Stable Diffusion (CompVis) ctx
—
Gemini 2.5 Pro ctx
1.0M
Capability radar
Value · context · multimodal · openness · speed posture
Value head-to-head
- Gemini 2.5 Pro93
Attribute tape
| Field | Stable Diffusion (CompVis) | Gemini 2.5 Pro |
|---|---|---|
| Developer | CompVis | Google DeepMind |
| Context | — | 1.0M |
| Modalities | Text, Code, Image, Audio, Video | |
| Openness | — | Proprietary Model |
| Speed | — | — |
| Price | — | $1.25/1M in · $10/1M out |
| Value score | — | 93 |
| Summary | The original open-source text-to-image latent diffusion model released by LMU Munich's CompVis lab in collaboration with Stability AI and Runway. It generates photorealistic images from text prompts and runs on consumer GPUs, with weights and code publicly available. | Google frontier multimodal model. |