all the models — AI benchmark observatory
← Models

Gemini 3.1 Flash-Lite

Google · proprietary · gemini-3.1-flash-lite-preview

Gemini 3.1 Flash-Lite is the first Flash-Lite model in the Gemini 3 series. It is optimized for high-volume, latency-sensitive tasks like translation, content moderation, and classification. It delivers enhanced performance at a fraction of the cost of larger models, with 2.5x faster Time to First Answer Token and 45% increased output speed compared to 2.5 Flash. Supports text, image, video, audio, and PDF input with a 1 million-token context window.

MMMLUGPQAMMMU-ProHumanity's L

Benchmark scores

BenchmarkScore
MMMLU0.89
GPQA0.87
MMMU-Pro0.77
Humanity's Last Exam0.16

Pricing

  • Google$0.25 / $1.50
  • DeepInfra$0.25 / $1.50

Input / output per 1M tokens

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.