all the models — AI benchmark observatory
← Models

Qwen3.6-35B-A3B

Alibaba Cloud / Qwen Team · open weight · qwen3.6-35b-a3b

Qwen3.6-35B-A3B is the first open-weight variant of the Qwen3.6 series, a multimodal Mixture-of-Experts model with 35B total parameters and 3B activated. It pairs a vision encoder with a hybrid 40-layer language model that interleaves Gated DeltaNet linear-attention blocks and Gated Attention blocks (10 × (3 × DeltaNet + 1 × Attention)) over 256 experts (8 routed + 1 shared, expert dim 512). The release prioritizes stability and real-world utility, with substantial gains in agentic coding (frontend workflows, repo-level reasoning) and a new option to preserve reasoning context across turns. Native context length is 262K tokens, extensible to ~1M via YaRN, and the model thinks by default.

MMLU-ReduxMMBench-V1.1AI2DAIME 2026RefCOCO-avgHMMT 2025

Benchmark scores

Pricing

  • DeepInfra$0.10 / $0.95

Input / output per 1M tokens

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.