all the models — AI benchmark observatory
← Benchmarks

IMO-AnswerBench

IMO-AnswerBench is a benchmark for evaluating mathematical reasoning capabilities on International Mathematical Olympiad (IMO) problems, focusing on answer generation and verification.

id imo-answerbench · max 1 · 22 models reported

#ModelScore
1Nemotron 3 Ultra (550B A55B)
NVIDIA · open
0.92
2GLM-5.2
Zhipu AI · open
0.91
3Hy3
Tencent · open
0.90
4Qwen3.7 Max
Alibaba Cloud / Qwen Team
0.90
5DeepSeek-V4-Pro-Max
DeepSeek · open
0.90