← Models
DeepSeek R1 Distill Llama 8B
DeepSeek · open weight · deepseek-r1-distill-llama-8b
DeepSeek-R1 is the first-generation reasoning model built atop DeepSeek-V3 (671B total parameters, 37B activated per token). It incorporates large-scale reinforcement learning (RL) to enhance its chain-of-thought and reasoning capabilities, delivering strong performance in math, code, and multi-step reasoning tasks.
Benchmark scores
| Benchmark | Score |
|---|---|
| MATH-500 | 0.89 |
| AIME 2024 | 0.80 |
| GPQA | 0.49 |
| LiveCodeBench | 0.40 |
Pricing
- No provider pricing.
AA metrics
No Artificial Analysis link yet.
Arena Elo
- No Arena snapshot linked.
