← Models
Gemini 3.1 Flash-Lite
Google · proprietary · gemini-3.1-flash-lite-preview
Gemini 3.1 Flash-Lite is the first Flash-Lite model in the Gemini 3 series. It is optimized for high-volume, latency-sensitive tasks like translation, content moderation, and classification. It delivers enhanced performance at a fraction of the cost of larger models, with 2.5x faster Time to First Answer Token and 45% increased output speed compared to 2.5 Flash. Supports text, image, video, audio, and PDF input with a 1 million-token context window.
Benchmark scores
| Benchmark | Score |
|---|---|
| MMMLU | 0.89 |
| GPQA | 0.87 |
| MMMU-Pro | 0.77 |
| Humanity's Last Exam | 0.16 |
Pricing
- Google$0.25 / $1.50
- DeepInfra$0.25 / $1.50
Input / output per 1M tokens
AA metrics
No Artificial Analysis link yet.
Arena Elo
- No Arena snapshot linked.
