← Models
Phi 4 Reasoning
Microsoft · open weight · phi-4-reasoning
Phi-4-reasoning is a state-of-the-art open-weight reasoning model finetuned from Phi-4 using supervised fine-tuning on a dataset of chain-of-thought traces and reinforcement learning. It focuses on math, science, and coding skills.
Benchmark scores
| Benchmark | Score |
|---|---|
| FlenQA | 0.98 |
| HumanEval+ | 0.93 |
| IFEval | 0.83 |
| AIME 2024 | 0.75 |
| MMLU-Pro | 0.74 |
| Arena Hard | 0.73 |
| GPQA | 0.66 |
| AIME 2025 | 0.63 |
| LiveCodeBench | 0.54 |
Pricing
- No provider pricing.
AA metrics
No Artificial Analysis link yet.
Arena Elo
- No Arena snapshot linked.
