← Models
Gemma 4 12B
Google · open weight · gemma-4-12b-it
Gemma 4 12B is Google DeepMind's encoder-free multimodal instruction-tuned model with 11.95 billion parameters and a 256K context window. It supports text, image, audio, and video inputs with text output, projecting image patches and audio waveforms directly into a single decoder-only transformer for streamlined local deployment.
Benchmark scores
| Benchmark | Score |
|---|---|
| GPQA | 0.79 |
| MMLU-Pro | 0.77 |
| MMMU-Pro | 0.69 |
| Humanity's Last Exam | 0.05 |
Pricing
- No provider pricing.
AA metrics
No Artificial Analysis link yet.
Arena Elo
- No Arena snapshot linked.
