all the models — AI benchmark observatory
← Models

Qwen2.5 32B Instruct

Alibaba Cloud / Qwen Team · open weight · qwen-2.5-32b-instruct

Qwen2.5-32B-Instruct is an instruction-tuned 32 billion parameter language model, part of the Qwen2.5 series. It is designed to follow instructions, generate long texts (over 8K tokens), understand structured data (e.g., tables), and generate structured outputs, especially JSON. The model supports multilingual capabilities across over 29 languages.

GSM8kHumanEvalMMLUMATHMMLU-ProGPQA

Benchmark scores

BenchmarkScore
GSM8k0.96
HumanEval0.88
MMLU0.83
MATH0.83
MMLU-Pro0.69
GPQA0.49

Pricing

  • No provider pricing.

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.
Qwen2.5 32B Instruct Benchmarks · all the models