← Models
Qwen3.5-4B
Alibaba Cloud / Qwen Team · open weight · qwen3.5-4b
Qwen3.5-4B is a 4 billion parameter vision-language model using Gated DeltaNet hybrid architecture with a 3:1 ratio of linear attention to full softmax attention. It supports 262K native context length and delivers strong performance for its size across knowledge, reasoning, coding, and multilingual tasks.
Benchmark scores
| Benchmark | Score |
|---|---|
| IFEval | 0.90 |
| MMLU-Redux | 0.89 |
| t2-bench | 0.80 |
| MMLU-Pro | 0.79 |
| GPQA | 0.76 |
Pricing
- No provider pricing.
AA metrics
No Artificial Analysis link yet.
Arena Elo
- No Arena snapshot linked.
