← Models
DeepSeek R1 Distill Llama 70B
DeepSeek · open weight · deepseek-r1-distill-llama-70b
DeepSeek-R1 is the first-generation reasoning model built atop DeepSeek-V3 (671B total parameters, 37B activated per token). It incorporates large-scale reinforcement learning (RL) to enhance its chain-of-thought and reasoning capabilities, delivering strong performance in math, code, and multi-step reasoning tasks.
Benchmark scores
| Benchmark | Score |
|---|---|
| MATH-500 | 0.94 |
| AIME 2024 | 0.87 |
| GPQA | 0.65 |
| LiveCodeBench | 0.57 |
Pricing
- No provider pricing.
AA metrics
No Artificial Analysis link yet.
Arena Elo
- No Arena snapshot linked.
