all the models — AI benchmark observatory
← Models

Claude Opus 4.7

Anthropic · proprietary · claude-opus-4-7

Claude Opus 4.7 is Anthropic's latest Opus-class model, a direct upgrade to Opus 4.6 with notable improvements in advanced software engineering, particularly on the most difficult tasks. It handles complex, long-running agentic workflows with rigor and consistency, follows instructions more literally and precisely, and verifies its own outputs before reporting back. Substantially improved vision supports high-resolution images up to 2,576 pixels on the long edge (~3.75 megapixels, over 3x prior Claude models), unlocking dense screenshot reading, complex diagram extraction, and pixel-perfect references. Better file system-based memory enables coherent multi-session work. Introduces a new 'xhigh' effort level between 'high' and 'max' for finer control over the reasoning/latency tradeoff, and ships with task budgets (public beta) on the Claude Platform. Uses an updated tokenizer (inputs may map to ~1.0-1.35x more tokens than Opus 4.6). Released with automated safeguards that detect and block prohibited or high-risk cybersecurity uses. Available across Claude products, the Claude API, Amazon Bedrock, Google Cloud Vertex AI, and Microsoft Foundry. Pricing: $5/$25 per million tokens (input/output), unchanged from Opus 4.6.

GPQAMMMLUCharXiv-RSWE-Bench VeBrowseCompOSWorld-Veri

Benchmark scores

Pricing

  • DeepInfra$5.00 / $25.00
  • Anthropic$5.00 / $25.00

Input / output per 1M tokens

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • text1494
  • vision1300
  • code1557