all the models — AI benchmark observatory
← Models

Claude Sonnet 4

Anthropic · proprietary · claude-sonnet-4-20250514

Claude Sonnet 4, part of the Claude 4 family, is a significant upgrade to Claude Sonnet 3.7. It excels in coding (72.7% on SWE-bench) and reasoning, responding more precisely to instructions. Sonnet 4 offers an optimal mix of capability and practicality, with enhanced steerability, and supports extended thinking with tool use.

GPQAMMMUSWE-Bench VeAIME 2025

Benchmark scores

BenchmarkScore
GPQA0.75
MMMU0.74
SWE-Bench Verified0.73
AIME 20250.70

Pricing

  • No provider pricing.

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.