all the models — AI benchmark observatory
← Models

GPT-5.2

OpenAI · proprietary · gpt-5.2-2025-12-11

GPT‑5.2 introduces substantial gains in professional knowledge work, outperforming experts on GDPval with 70.9% wins or ties, and setting new highs in coding (SWE‑Bench Pro 55.6%), science (GPQA Diamond ~92–93%), math (AIME 2025: 100%), long‑context accuracy up to 256k tokens, and reliable tool‑calling (Tau2 Telecom 98.7%). It rolls out as Instant, Thinking, and Pro—faster, more structured, and less error‑prone—priced at $1.75/1M input and $14/1M output tokens, with Pro variants supporting xhigh reasoning for top‑quality, end‑to‑end execution.

AIME 2025HMMT 2025Tau2 TelecomGraphwalks BGPQAMMMLU

Benchmark scores

Pricing

  • OpenAI$1.75 / $14.00

Input / output per 1M tokens

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.
GPT-5.2 Benchmarks · all the models