all the models — AI benchmark observatory
← Models

Qwen3-235B-A22B-Instruct-2507

Alibaba Cloud / Qwen Team · open weight · qwen3-235b-a22b-instruct-2507

Qwen3-235B-A22B-Instruct-2507 is the updated instruct version of Qwen3-235B-A22B featuring significant improvements in general capabilities including instruction following, logical reasoning, text comprehension, mathematics, science, coding and tool usage. It provides substantial gains in long-tail knowledge coverage across multiple languages and markedly better alignment with user preferences in subjective and open-ended tasks.

ZebraLogicMMLU-ReduxIFEvalMMLU-ProGPQABFCL-v3

Benchmark scores

BenchmarkScore
ZebraLogic0.95
MMLU-Redux0.93
IFEval0.89
MMLU-Pro0.83
GPQA0.78
BFCL-v30.71
AIME 20250.70

Pricing

  • DeepInfra$0.09 / $0.55

Input / output per 1M tokens

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.