all the models — AI benchmark observatory
← Models

Qwen2.5 7B Instruct

Alibaba Cloud / Qwen Team · open weight · qwen-2.5-7b-instruct

Qwen2.5-7B-Instruct is an instruction-tuned 7B parameter language model that excels at following instructions, generating long texts (over 8K tokens), understanding structured data, and generating structured outputs like JSON. The model features enhanced capabilities in mathematics, coding, and multilingual support across 29+ languages including Chinese, English, French, Spanish, and more.

GSM8kHumanEvalMATHIFEvalMMLU-ProArena Hard

Benchmark scores

BenchmarkScore
GSM8k0.92
HumanEval0.85
MATH0.76
IFEval0.71
MMLU-Pro0.56
Arena Hard0.52
GPQA0.36
LiveCodeBench0.29

Pricing

  • No provider pricing.

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.