DeepSeek-V4-Flash-Max
DeepSeek · open weight · deepseek-v4-flash-max
DeepSeek-V4-Flash-Max is the maximum reasoning effort mode of DeepSeek-V4-Flash, a 284B-parameter MoE model with 13B activated parameters and a 1M-token context window. Sharing the V4 series' hybrid attention architecture (Compressed Sparse Attention combined with Heavily Compressed Attention), Manifold-Constrained Hyper-Connections, and Muon optimizer, V4-Flash-Max delivers reasoning performance comparable to V4-Pro when given a larger thinking budget while operating at a fraction of the parameter scale. It is pre-trained on more than 32T tokens and post-trained with a two-stage paradigm of domain-specific expert cultivation followed by on-policy distillation.
Benchmark scores
| Benchmark | Score |
|---|---|
| CodeForces | 1.00 |
| HMMT Feb 26 | 0.95 |
| LiveCodeBench | 0.92 |
| GPQA | 0.88 |
| MMLU-Pro | 0.86 |
| SWE-Bench Verified | 0.79 |
| SWE-bench Multilingual | 0.73 |
| BrowseComp | 0.73 |
| Humanity's Last Exam | 0.45 |
Pricing
- DeepSeek$0.14 / $0.28
- DeepInfra$0.09 / $0.18
Input / output per 1M tokens
AA metrics
No Artificial Analysis link yet.
Arena Elo
- No Arena snapshot linked.
