Model comparison
CLAP HTSAT-Fused vs StarCoder2 15B
Capability radar, value bars, and side-by-side posture.
CLAP HTSAT-Fused
Not scored
value score
StarCoder2 15B
70
value score
CLAP HTSAT-Fused ctx
—
StarCoder2 15B ctx
16K
Capability radar
Value · context · multimodal · openness · speed posture
Value head-to-head
- StarCoder2 15B70
Attribute tape
| Field | CLAP HTSAT-Fused | StarCoder2 15B |
|---|---|---|
| Developer | Hugging Face | Hugging Face |
| Context | — | 16K |
| Modalities | Audio-classification | Text, Code |
| Openness | — | Open Weights |
| Speed | — | — |
| Price | — | — |
| Value score | — | 70 |
| Summary | CLAP HTSAT-Fused is an open-source contrastive audio-language pretraining model from LAION that maps audio clips and text prompts into a shared embedding space. It supports zero-shot audio classification, audio-text retrieval, and feature extraction, using feature fusion to handle variable-length audio inputs. | BigCode open coding model on Hugging Face. |