all the models — AI benchmark observatory
← Models

Gemma 3 4B

Google · open weight · gemma-3-4b-it

Gemma 3 4B is a 4-billion-parameter vision-language model from Google, handling text and image input and generating text output. It features a 128K context window, multilingual support, and open weights. Suitable for question answering, summarization, reasoning, and image understanding tasks.

IFEvalGSM8kMATHHumanEvalMMLU-ProGPQA

Benchmark scores

BenchmarkScore
IFEval0.90
GSM8k0.89
MATH0.76
HumanEval0.71
MMLU-Pro0.44
GPQA0.31
LiveCodeBench0.13

Pricing

  • DeepInfra$0.05 / $0.10

Input / output per 1M tokens

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.