all the models — AI benchmark observatory
← Models

Claude 3.7 Sonnet

Anthropic · proprietary · claude-3-7-sonnet-20250219

The most intelligent Claude model and the first hybrid reasoning model on the market. Claude 3.7 Sonnet can produce near-instant responses or extended, step-by-step thinking that is made visible to the user. Shows particularly strong improvements in coding and front-end web development.

MATH-500IFEvalGPQAAIME 2024MMMUSWE-Bench Ve

Benchmark scores

BenchmarkScore
MATH-5000.96
IFEval0.93
GPQA0.85
AIME 20240.80
MMMU0.75
SWE-Bench Verified0.70
AIME 20250.55

Pricing

  • No provider pricing.

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.