← Benchmarks
We-Math
We-Math evaluates multimodal models on visual mathematical reasoning, requiring models to understand and solve math problems presented with visual elements such as diagrams, charts, and geometric figures.
id we-math · max 1 · 1 models reported
| # | Model | Score |
|---|---|---|
| 1 | Qwen3.6 Plus Alibaba Cloud / Qwen Team | 0.89 |
