all the models — AI benchmark observatory
← Models

Llama 3.2 11B Instruct

Meta · open weight · llama-3.2-11b-instruct

Llama 3.2 11B Vision Instruct is an instruction-tuned multimodal large language model optimized for visual recognition, image reasoning, captioning, and answering general questions about an image. It accepts text and images as input and generates text as output.

AI2DDocVQAMMLUMATHMMMUMMMU-Pro

Benchmark scores

BenchmarkScore
AI2D0.91
DocVQA0.88
MMLU0.73
MATH0.52
MMMU0.51
MMMU-Pro0.33
GPQA0.33

Pricing

  • No provider pricing.

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.
Llama 3.2 11B Instruct Benchmarks · all the models