Model Terminal
DeepSeek-V4-Flash-0731-CB-C16-NVFP4 (JasonW2025 NVFP4 build)
A third-party NVIDIA FP4-quantized redistribution of DeepSeek's open-weight DeepSeek-V4-Flash-0731 model, hosted by the Hugging Face user JasonW2025. The base model is a 284B-parameter Mixture-of-Experts LLM (13B active parameters) optimized for reasoning, coding, agentic workflows, and long-context chat up to approximately 1M tokens. Source: supabase.
Value score
Not scored
Context
—
tokens
Max output
—
tokens
Price
—
Capability radar
Peer value bars
Identity
- Developer
- JasonW2025
- Openness
- —
- Modalities
- —
- Release
- —
- Knowledge cutoff
- —
- Deprecation
- —
- API docs
- —
Benchmarks
No benchmark scores yet.
Compare nearby
Pricing
- Input / 1M
- —
- Output / 1M
- —
- Speed
- —