all the models — AI benchmark observatory
← Models

Mistral Small 3 24B Instruct

Mistral AI · open weight · mistral-small-24b-instruct-2501

Mistral Small 3 is a 24B-parameter LLM licensed under Apache-2.0. It focuses on low-latency, high-efficiency instruction following, maintaining performance comparable to larger models. It provides quick, accurate responses for conversational agents, function calling, and domain-specific fine-tuning. Suitable for local inference when quantized, it rivals models 2–3× its size while using significantly fewer compute resources.

Arena HardHumanEvalIFEvalMATHMMLU-ProGPQA

Benchmark scores

BenchmarkScore
Arena Hard0.88
HumanEval0.85
IFEval0.83
MATH0.71
MMLU-Pro0.66
GPQA0.45

Pricing

  • DeepInfra$0.05 / $0.08

Input / output per 1M tokens

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.
Mistral Small 3 24B Instruct Benchmarks · all the models