Model Terminal
NVIDIA Nemotron 3 Super
Nemotron 3 Super is NVIDIA's open-weight reasoning LLM for agentic workflows, featuring a hybrid Mamba-2 + MoE + attention architecture with approximately 120B total and 12B active parameters per token. It is designed for efficient long-context inference, tool use, and multi-step reasoning in coding, RAG, and IT automation tasks. Source: supabase.
Value score
Not scored
Context
—
tokens
Max output
—
tokens
Price
—
Capability radar
Peer value bars
Identity
- Developer
- NVIDIA
- Openness
- —
- Modalities
- —
- Release
- —
- Knowledge cutoff
- —
- Deprecation
- —
- API docs
- —
Benchmarks
No benchmark scores yet.
Compare nearby
Pricing
- Input / 1M
- —
- Output / 1M
- —
- Speed
- —