all the models — AI benchmark observatory
← Benchmarks

QwenWebBench

QwenWebBench is an internal front-end code generation benchmark by Qwen. It is bilingual (EN/CN) and spans 7 categories (Web Design, Web Apps, Games, SVG, Data Visualization, Animation, and 3D), using auto-render plus a multimodal judge for code and visual correctness. Scores are reported as BT/Elo ratings.

id qwenwebbench · max 2000 · 2 models reported

#ModelScore
1Qwen3.7 Max
Alibaba Cloud / Qwen Team
1568.00
2Qwen3.6-27B
Alibaba Cloud / Qwen Team · open
1487.00
QwenWebBench Leaderboard · all the models