Model Terminal

GLM-4.5V

GLM-4.5V is an open-weight multimodal vision-language model from Z.ai featuring a Mixture-of-Experts architecture (106B total / 12B active parameters). It is designed for image reasoning, document parsing, video understanding, GUI screen reading, and visual grounding tasks. Source: supabase.

Value score
Not scored
Context
tokens
Max output
tokens
Price

Capability radar

Identity

Developer
Z.ai
Openness
Modalities
Release
Knowledge cutoff
Deprecation
API docs

Benchmarks

No benchmark scores yet.

Compare nearby

Pricing

Input / 1M
Output / 1M
Speed