all the models — AI benchmark observatory
← Models

o1-preview

OpenAI · proprietary · o1-preview

A research preview model focused on mathematical and logical reasoning capabilities, demonstrating improved performance on tasks requiring step-by-step reasoning, mathematical problem-solving, and code generation. The model shows enhanced capabilities in formal reasoning while maintaining strong general capabilities.

MGSMMMLUMATHGPQAAIME 2024SWE-Bench Ve

Benchmark scores

BenchmarkScore
MGSM0.91
MMLU0.91
MATH0.85
GPQA0.73
AIME 20240.42
SWE-Bench Verified0.41

Pricing

  • No provider pricing.

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.
o1-preview Benchmarks · all the models