all the models — AI benchmark observatory
← Models

Qwen3.8 Flash

Alibaba Cloud / Qwen Team · proprietary · qwen3.8-flash

Qwen3.8 Flash is the production QwenCloud / OpenRouter API model (id qwen3.8-flash), not the open-weight Qwen3.8-Flash-Next checkpoint. Hugging Face states Flash is the official managed version based on Flash-Next with production features such as 1M context by default and official built-in tools. Architecture inherits Flash-Next (125B parameters with 6B activated, plus 51B n-gram embedding and 4B MTP). Multimodal text, image, and video understanding; thinking on by default; function calling, built-in tools, and structured output. QwenCloud lists 1M context, ~128k max output, and ~256k thinking budget. Open weights remain under the separate catalog card qwen3.8-flash-next.

MathVisionGPQACharXiv-RRealWorldQAAndroidWorldSWE-bench Mu

Benchmark scores

Pricing

  • Novita$0.15 / $0.47

Input / output per 1M tokens

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.