all the models — AI benchmark observatory
← Models

GPT OSS 120B

OpenAI · open weight · gpt-oss-120b

GPT-OSS-120B is an open-weight, 116.8B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized to run on a single H100 GPU with native MXFP4 quantization. The model supports configurable reasoning depth, full chain-of-thought access, and native tool use, including function calling, browsing, and structured output generation. It achieves near-parity with OpenAI o4-mini on core reasoning benchmarks. Note: While referred to as '120b' for simplicity, it technically has 116.8B parameters.

MMLUGPQAHumanity's L

Benchmark scores

BenchmarkScore
MMLU0.90
GPQA0.80
Humanity's Last Exam0.15

Pricing

  • DeepInfra$0.04 / $0.17

Input / output per 1M tokens

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.