← Models
GPT-5.2
OpenAI · proprietary · gpt-5.2-2025-12-11
GPT‑5.2 introduces substantial gains in professional knowledge work, outperforming experts on GDPval with 70.9% wins or ties, and setting new highs in coding (SWE‑Bench Pro 55.6%), science (GPQA Diamond ~92–93%), math (AIME 2025: 100%), long‑context accuracy up to 256k tokens, and reliable tool‑calling (Tau2 Telecom 98.7%). It rolls out as Instant, Thinking, and Pro—faster, more structured, and less error‑prone—priced at $1.75/1M input and $14/1M output tokens, with Pro variants supporting xhigh reasoning for top‑quality, end‑to‑end execution.
Benchmark scores
| Benchmark | Score |
|---|---|
| AIME 2025 | 1.00 |
| HMMT 2025 | 0.99 |
| Tau2 Telecom | 0.99 |
| Graphwalks BFS <128k | 0.94 |
| GPQA | 0.92 |
| MMMLU | 0.90 |
| ScreenSpot Pro | 0.86 |
| ARC-AGI | 0.86 |
| VideoMMMU | 0.86 |
| SWE-Bench Verified | 0.80 |
| MMMU-Pro | 0.80 |
| SWE-Lancer (IC-Diamond subset) | 0.75 |
| Humanity's Last Exam | 0.34 |
Pricing
- OpenAI$1.75 / $14.00
Input / output per 1M tokens
AA metrics
No Artificial Analysis link yet.
Arena Elo
- No Arena snapshot linked.
