all the models — AI benchmark observatory
← Benchmarks

HorizonMath

HorizonMath is an extremely difficult frontier mathematics benchmark designed to test the limits of mathematical reasoning on research-level and competition-beyond problems.

id horizonmath · max 1 · 4 models reported

#ModelScore

No scores for this benchmark yet.