Qwen3.6-35B-A3B
Alibaba Cloud / Qwen Team · open weight · qwen3.6-35b-a3b
Qwen3.6-35B-A3B is the first open-weight variant of the Qwen3.6 series, a multimodal Mixture-of-Experts model with 35B total parameters and 3B activated. It pairs a vision encoder with a hybrid 40-layer language model that interleaves Gated DeltaNet linear-attention blocks and Gated Attention blocks (10 × (3 × DeltaNet + 1 × Attention)) over 256 experts (8 routed + 1 shared, expert dim 512). The release prioritizes stability and real-world utility, with substantial gains in agentic coding (frontend workflows, repo-level reasoning) and a new option to preserve reasoning context across turns. Native context length is 262K tokens, extensible to ~1M via YaRN, and the model thinks by default.
Benchmark scores
| Benchmark | Score |
|---|---|
| MMLU-Redux | 0.93 |
| MMBench-V1.1 | 0.93 |
| AI2D | 0.93 |
| AIME 2026 | 0.93 |
| RefCOCO-avg | 0.92 |
| HMMT 2025 | 0.91 |
| OmniDocBench 1.5 | 0.90 |
| HMMT25 | 0.89 |
| VideoMME w sub. | 0.87 |
| MathVista-Mini | 0.86 |
| MLVU | 0.86 |
| GPQA | 0.86 |
| RealWorldQA | 0.85 |
| MMLU-Pro | 0.85 |
| MMMU | 0.82 |
| MMMU-Pro | 0.75 |
| SWE-Bench Verified | 0.73 |
| Humanity's Last Exam | 0.21 |
Pricing
- DeepInfra$0.10 / $0.95
Input / output per 1M tokens
AA metrics
No Artificial Analysis link yet.
Arena Elo
- No Arena snapshot linked.
