← Models
Claude Opus 4.6
Anthropic · proprietary · claude-opus-4-6
Claude Opus 4.6 is Anthropic's most intelligent model, improving on its predecessor's coding skills with more careful planning, longer agentic task sustenance, more reliable operation in larger codebases, and better code review and debugging skills. First Opus-class model with 1M token context window (beta), 128K output tokens, and adaptive thinking. Features effort controls (low/medium/high/max) and context compaction for long-running tasks. State-of-the-art on Terminal-Bench 2.0, Humanity's Last Exam, GDPval-AA, and BrowseComp. Pricing: $5/$25 per million tokens (input/output).
Benchmark scores
| Benchmark | Score |
|---|---|
| Vending-Bench 2 | 8017.59 |
| AIME 2025 | 1.00 |
| Tau2 Telecom | 0.99 |
| Graphwalks parents >128k | 0.95 |
| DeepSearchQA | 0.91 |
| GPQA | 0.91 |
| MMMLU | 0.91 |
| BrowseComp | 0.84 |
| SWE-Bench Verified | 0.81 |
| SWE-bench Multilingual | 0.78 |
| MMMU-Pro | 0.77 |
| CyberGym | 0.74 |
| OSWorld | 0.73 |
| Humanity's Last Exam | 0.53 |
Pricing
- Anthropic$5.00 / $25.00
Input / output per 1M tokens
AA metrics
No Artificial Analysis link yet.
Arena Elo
- text1497
- vision1293
- code1537
