all the models — AI benchmark observatory
← Models

Qwen2.5-Coder 7B Instruct

Alibaba Cloud / Qwen Team · open weight · qwen-2.5-coder-7b-instruct

Qwen2.5-Coder is a specialized coding model trained on 5.5 trillion tokens of code data, supporting 92 programming languages with a 128K context window. It excels in code generation, completion, and repair while maintaining strong performance in math and general tasks. The model demonstrates exceptional capabilities in multi-programming language tasks and code reasoning.

HumanEvalMMLUMATHMMLU-ProLiveCodeBenc

Benchmark scores

BenchmarkScore
HumanEval0.88
MMLU0.68
MATH0.47
MMLU-Pro0.40
LiveCodeBench0.18

Pricing

  • No provider pricing.

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.
Qwen2.5-Coder 7B Instruct Benchmarks · all the models