all the models — AI benchmark observatory
← Models

Kimi K2 Base

Moonshot AI · open weight · kimi-k2-base

Kimi K2 base model is a state-of-the-art mixture-of-experts (MoE) language model with 32 billion activated parameters and 1 trillion total parameters. Trained on 15.5 trillion tokens with the MuonClip optimizer, this is the foundation model before instruction tuning. It demonstrates strong performance on knowledge, reasoning, and coding benchmarks while being optimized for agentic capabilities.

C-EvalGSM8kMMLU-redux-2MMLUEvalPlusMATH

Benchmark scores

BenchmarkScore
C-Eval0.93
GSM8k0.92
MMLU-redux-2.00.90
MMLU0.88
EvalPlus0.80
MATH0.70
MMLU-Pro0.69
GPQA0.48

Pricing

  • No provider pricing.

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.