all the models — AI benchmark observatory
← Models

Qwen2.5-Coder 32B Instruct

Alibaba Cloud / Qwen Team · open weight · qwen-2.5-coder-32b-instruct

Qwen2.5-Coder is a specialized coding model trained on 5.5 trillion tokens of code data, supporting 92 programming languages with a 128K context window. It excels in code generation, completion, repair, and multi-programming tasks while maintaining strong performance in mathematics and general capabilities.

HumanEvalGSM8kMMLUMATHMMLU-ProLiveCodeBenc

Benchmark scores

BenchmarkScore
HumanEval0.93
GSM8k0.91
MMLU0.75
MATH0.57
MMLU-Pro0.50
LiveCodeBench0.31

Pricing

  • No provider pricing.

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.