all the models — AI benchmark observatory
← Models

Claude 3 Opus

Anthropic · proprietary · claude-3-opus-20240229

Claude 3 Opus is Anthropic's most intelligent model, with best-in-market performance on highly complex tasks. It can navigate open-ended prompts and sight-unseen scenarios with remarkable fluency and human-like understanding, showing the outer limits of what's possible with generative AI.

ARC-CHellaSwagGSM8kMGSMMMLUHumanEval

Benchmark scores

BenchmarkScore
ARC-C0.96
HellaSwag0.95
GSM8k0.95
MGSM0.91
MMLU0.87
HumanEval0.85
MMLU-Pro0.69
MATH0.60
GPQA0.50

Pricing

  • No provider pricing.

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.