Model comparison
UniDepthV2 ViT-L/14 vs StarCoder2 15B
Capability radar, value bars, and side-by-side posture.
UniDepthV2 ViT-L/14
Not scored
value score
StarCoder2 15B
70
value score
UniDepthV2 ViT-L/14 ctx
—
StarCoder2 15B ctx
16K
Capability radar
Value · context · multimodal · openness · speed posture
Value head-to-head
- StarCoder2 15B70
Attribute tape
| Field | UniDepthV2 ViT-L/14 | StarCoder2 15B |
|---|---|---|
| Developer | Hugging Face | Hugging Face |
| Context | — | 16K |
| Modalities | Text | Text, Code |
| Openness | — | Open Weights |
| Speed | — | — |
| Price | — | — |
| Value score | — | 70 |
| Summary | UniDepthV2 ViT-L/14 is a large-scale monocular metric depth estimation model that predicts real-world depth from a single RGB image without requiring camera intrinsics at inference time. It is the highest-capacity checkpoint in the UniDepthV2 family, offering metric 3D scene estimates with per-pixel confidence outputs across diverse camera types and scenes. | BigCode open coding model on Hugging Face. |