all the models — AI benchmark observatory
← Models

Grok 4.7

xAI · proprietary · grok-4.7

Grok 4.7 is SpaceXAI's frontier model for coding, agentic tasks, and knowledge work. It uses a larger base model than Grok 4.6, with longer RL on harder multi-hour tasks, stronger self-verification, and native Grok Bot harness understanding. Reasoning efforts: low, medium, high, xhigh (default high). 500K context; same price and speed class as Grok 4.6 ($2/$0.50 cached/$6 per 1M tokens under 200k prompts; ≥200k prompts bill at $4/$1/$12). Grok 4.7 Fast uses the same model on faster infrastructure at twice the token rates and twice the output speed, and is available only in Cursor and Grok Build — not as a separate public API model id (docs expose only `grok-4.7`). Pretraining knowledge cutoff June 2026, with supplemental training data through August 2026. Multimodal: text and image in, text out.

GDPval (Elo)AA-BriefcaseCyberGym

Benchmark scores

BenchmarkScore
GDPval (Elo)1695.00
AA-Briefcase v1.11657.00
CyberGym0.80

Pricing

  • xAI$2.00 / $6.00

Input / output per 1M tokens

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.
Grok 4.7 Benchmarks · all the models