all the models — AI benchmark observatory
← Models

GPT-3.5 Turbo

OpenAI · proprietary · gpt-3.5-turbo-0125

The latest GPT-3.5 Turbo model with higher accuracy at responding in requested formats and a fix for a bug which caused a text encoding issue for non-English language function calls.

MMLUHumanEvalMATHGPQA

Benchmark scores

BenchmarkScore
MMLU0.70
HumanEval0.68
MATH0.43
GPQA0.31
MMMU0.00

Pricing

  • OpenAI$0.50 / $1.50
  • Azure$0.50 / $1.50

Input / output per 1M tokens

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.