all the models — AI benchmark observatory
← Models

Granite 3.3 8B Instruct

IBM · open weight · granite-3.3-8b-instruct

Granite 3.3 models feature enhanced reasoning capabilities and support for Fill-in-the-Middle (FIM) code completion. They are built on a foundation of open-source instruction datasets with permissive licenses, alongside internally curated synthetic datasets tailored for long-context problem-solving. These models preserve the key strengths of previous Granite versions, including support for a 128K context length, strong performance in retrieval-augmented generation (RAG) and function calling, and controls for response length and originality. Granite 3.3 also delivers competitive results across general, enterprise, and safety benchmarks. Released as open source, the models are available under the Apache 2.0 license.

HumanEvalAIME 2024IFEvalMMLUArena Hard

Benchmark scores

BenchmarkScore
HumanEval0.90
AIME 20240.81
IFEval0.75
MMLU0.66
Arena Hard0.58

Pricing

  • No provider pricing.

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.
Granite 3.3 8B Instruct Benchmarks · all the models