all the models — AI benchmark observatory
← Models

Grok-1.5

xAI · proprietary · grok-1.5

An advanced language model with improved reasoning capabilities, particularly excelling in coding and mathematical tasks. Features a 128K token context window and enhanced problem-solving abilities compared to its predecessor.

GSM8kDocVQAMMLUHumanEvalMMMUMMLU-Pro

Benchmark scores

BenchmarkScore
GSM8k0.90
DocVQA0.86
MMLU0.81
HumanEval0.74
MMMU0.54
MMLU-Pro0.51
MATH0.51
GPQA0.36

Pricing

  • No provider pricing.

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.