all the models — AI benchmark observatory
← Models

Phi-3.5-MoE-instruct

Microsoft · open weight · phi-3.5-moe-instruct

Phi-3.5-MoE-instruct is a mixture-of-experts model with ~42B total parameters (6.6B active) and a 128K context window. It excels at reasoning, math, coding, and multilingual tasks, outperforming larger dense models in many benchmarks. It underwent a thorough safety post-training process (SFT + DPO) and is licensed under MIT. This model is ideal for scenarios where efficiency and high performance are both required, particularly in multi-lingual or reasoning-intensive tasks.

GSM8kRepoQAMMLUHumanEvalMATHMMLU-Pro

Benchmark scores

BenchmarkScore
GSM8k0.89
RepoQA0.85
MMLU0.79
HumanEval0.71
MATH0.59
MMLU-Pro0.45
Arena Hard0.38
GPQA0.37

Pricing

  • No provider pricing.

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.