← Models
Llama 3.1 Nemotron Ultra 253B v1
NVIDIA · open weight · llama-3.1-nemotron-ultra-253b-v1
A 253B parameter derivative of Meta Llama 3.1 405B Instruct, developed by NVIDIA using Neural Architecture Search (NAS) and vertical compression. It underwent multi-phase post-training (SFT for Math, Code, Reasoning, Chat, Tool Calling; RL with GRPO) to enhance reasoning and instruction-following. Optimized for accuracy/efficiency tradeoff on NVIDIA GPUs. Supports 128k context.
Benchmark scores
| Benchmark | Score |
|---|---|
| MATH-500 | 0.97 |
| IFEval | 0.89 |
| GPQA | 0.76 |
| AIME 2025 | 0.72 |
| LiveCodeBench | 0.66 |
Pricing
- No provider pricing.
AA metrics
No Artificial Analysis link yet.
Arena Elo
- No Arena snapshot linked.
