Available in Shortly AI
Gemini 3.1 Flash Lite
Released May 7, 2026
ReasoningVision (Image)Tool Use
Gemini 3.1 Flash Lite outperforms 2.5 Flash Lite on overall quality and lands close to 2.5 Flash performance across key capability areas. It is a workhorse model for high-volume use cases, with improvements across audio input/ASR, RAG snippet ranking, translation, data extraction, and code completion.
- Context window
- 1M tokens
- Maximum output
- 65K tokens
- Input pricing
- $0.25/M
- Output pricing
- $1.50/M
Independent benchmark
LiveBench historical snapshot
61.7overall
Historical benchmark data from LiveBench, release 2026-01-08. Scores are pinned and are not fetched on page load.
59.7
68.5
33.3
73.6
54.9
73.2
68.6
| Category | Score |
|---|---|
| Reasoning | 59.659 |
| Coding | 68.524 |
| Agentic Coding | 33.333 |
| Mathematics | 73.559 |
| Data Analysis | 54.897 |
| Language | 73.183 |
| Instruction Following | 68.617 |
Benchmarks measure selected evaluation tasks and do not guarantee performance for every prompt or workflow.
FAQ
Similar AI models
Inkling SmallThinkingmachinesInkling-Small is a lighter-weight model with 12B active parameters, trained with a similar recipe, to Inkling that achieves strong performance with even lower cost and latency.Gemini 3.5 Flash LiteGoogleGemini 3.5 Flash Lite features upgraded agentic capabilities, making the model ideal for subagents in complex workflows.Gemini 3.6 FlashGoogleGemini 3.6 Flash delivers higher quality across coding, agentic workflows, and web development with reduced token consumption and fewer model calls compared to previous model iterations.
Available with one subscription
Start with Gemini 3.1 Flash Lite in Shortly AI
Keep your conversations, files and model choices together in one workspace.