all the models — AI benchmark observatory
← Models

Claude Opus 4.5

Anthropic · proprietary · claude-opus-4-5-20251101

Premium model combining maximum intelligence with practical performance. Best model in the world for coding, agents, and computer use. Most robustly aligned model with best prompt injection resistance of any frontier model. Features extended thinking, 200K context window, 64K max output, and a new effort parameter for controlling reasoning depth. Pricing: $5/$25 per million tokens (input/output).

Tau2 TelecomMMMLUGPQASWE-Bench Ve

Benchmark scores

BenchmarkScore
Tau2 Telecom0.98
MMMLU0.91
GPQA0.87
SWE-Bench Verified0.81

Pricing

  • No provider pricing.

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • text1469
  • code1468
Claude Opus 4.5 Benchmarks · all the models