all the models — AI benchmark observatory
← Models

Qwen3.7 Max

Alibaba Cloud / Qwen Team · proprietary · qwen3.7-max

Qwen3.7 Max is Alibaba Cloud Qwen Team's proprietary flagship model for agent-driven workflows. It is designed for coding agents, office automation, MCP and multi-agent orchestration, and long-horizon autonomous execution, with a 1 million token context window and up to 65,536 output tokens. Qwen reports strong agentic coding results including 69.7 on Terminal-Bench 2.0-Terminus, 80.4 on SWE-bench Verified, 60.6 on SWE-Pro, and 78.3 on SWE-Multilingual, alongside 92.4 on GPQA Diamond and 97.1 on HMMT 2026 Feb.

QwenSVGQwenWebBenchHMMT Feb 26Kernel BenchMMLU-ReduxIFEval

Benchmark scores

Pricing

  • Novita$1.25 / $3.75
  • Together$2.50 / $7.50
  • DeepInfra$2.50 / $7.50

Input / output per 1M tokens

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.