all the models — AI benchmark observatory
← Models

Qwen3 235B A22B

Alibaba Cloud / Qwen Team · open weight · qwen3-235b-a22b

Qwen3 235B A22B is a large language model developed by Alibaba, featuring a Mixture-of-Experts (MoE) architecture with 235 billion total parameters and 22 billion activated parameters. It achieves competitive results in benchmark evaluations of coding, math, general capabilities, and more, compared to other top-tier models.

Arena HardGSM8kBBHMMLUAIME 2024AIME 2025

Benchmark scores

BenchmarkScore
Arena Hard0.96
GSM8k0.94
BBH0.89
MMLU0.88
AIME 20240.86
AIME 20250.81
EvalPlus0.78
MATH0.72
BFCL0.71
LiveCodeBench0.71
MMLU-Pro0.68
GPQA0.47

Pricing

  • No provider pricing.

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.