all the models — AI benchmark observatory
← Models

Claude Opus 4

Anthropic · proprietary · claude-opus-4-20250514

Claude Opus 4 is Anthropic's most powerful model and the world's best coding model, part of the Claude 4 family. It delivers sustained performance on complex, long-running tasks and agent workflows. Opus 4 excels at coding, advanced reasoning, and can use tools (like web search) during extended thinking. It supports parallel tool execution and has improved memory capabilities.

MMMLUGPQAAIME 2025SWE-Bench Ve

Benchmark scores

BenchmarkScore
MMMLU0.89
GPQA0.80
AIME 20250.76
SWE-Bench Verified0.72

Pricing

  • No provider pricing.

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.