all the models — AI benchmark observatory
← Models

GPT-5.3 Codex

OpenAI · proprietary · gpt-5.3-codex

GPT-5.3-Codex is OpenAI's most capable coding model, combining frontier agentic coding capabilities, improvements in aesthetics, and context compaction. It sets new state-of-the-art results on Terminal-Bench 2.0 (77.3%), OSWorld-Verified (64.7%), and SWE-Lancer IC Diamond (81.4%). First model classified as High capability for cybersecurity under OpenAI's Preparedness Framework. Available in the Codex app and API.

Not enough scores for a fingerprint yet.

Benchmark scores

Pricing

  • OpenAI$1.75 / $14.00

Input / output per 1M tokens

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.
GPT-5.3 Codex Benchmarks · all the models