all the models — AI benchmark observatory
← Models

QwQ-32B

Alibaba Cloud / Qwen Team · open weight · qwq-32b

A model focused on advancing AI reasoning capabilities, particularly excelling in mathematics and programming. Features deep introspection and self-questioning abilities while having some limitations in language mixing and recursive/endless reasoning patterns.

MATH-500IFEvalAIME 2024BFCLGPQALiveCodeBenc

Benchmark scores

BenchmarkScore
MATH-5000.91
IFEval0.84
AIME 20240.80
BFCL0.66
GPQA0.65
LiveCodeBench0.63

Pricing

  • No provider pricing.

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.
QwQ-32B Benchmarks · all the models