Model Terminal

Qwen3.6-35B-A3B-GGUF (Unsloth)

Unsloth's GGUF-quantized packaging of Alibaba's Qwen3.6-35B-A3B sparse mixture-of-experts model, enabling local and self-hosted inference on consumer-grade hardware via llama.cpp and compatible runtimes. The model features 35B total parameters with only 3B active per token and a 262K-token native context window. Source: supabase.

Value score
Not scored
Context
tokens
Max output
tokens
Price

Capability radar

Identity

Developer
unsloth
Openness
Modalities
Release
Knowledge cutoff
Deprecation
API docs

Benchmarks

No benchmark scores yet.

Compare nearby

Pricing

Input / 1M
Output / 1M
Speed