Model Terminal

DeepSeek-R1-Distill-Qwen-7B

A 7B open-weight reasoning model distilled from DeepSeek-R1, built on the Qwen2.5-Math-7B base. It preserves much of DeepSeek-R1's step-by-step reasoning capability at a fraction of the size, making it practical for local deployment and low-cost inference on math, code, and general reasoning tasks. Source: supabase.

Value score
Not scored
Context
128K
tokens
Max output
tokens
Price

Capability radar

Identity

Developer
deepseek-ai
Openness
Open Weights
Modalities
Release
2025-01-20
Knowledge cutoff
Deprecation

Benchmarks

  • GPQA71.5 · 2025-01-20 · source

Compare nearby

Pricing

Input / 1M
Output / 1M
Speed