Available in Shortly AI
Gemini 3.5 Flash Lite
Released July 21, 2026
ReasoningVision (Image)Tool Use
Gemini 3.5 Flash Lite features upgraded agentic capabilities, making the model ideal for subagents in complex workflows.
- Context window
- 1M tokens
- Maximum output
- 65K tokens
- Input pricing
- $0.30/M
- Output pricing
- $2.50/M
Independent benchmark
LiveBench performance
63.9overall
Benchmark data from LiveBench, release 2026-06-25. Scores are pinned and are not fetched on page load.
60.2
76.1
45.3
73.7
53.3
71.8
67.2
| Category | Score |
|---|---|
| Reasoning | 60.188 |
| Coding | 76.072 |
| Agentic Coding | 45.253 |
| Mathematics | 73.739 |
| Data Analysis | 53.250 |
| Language | 71.820 |
| Instruction Following | 67.238 |
Benchmarks measure selected evaluation tasks and do not guarantee performance for every prompt or workflow.
FAQ
Similar AI models
Inkling SmallThinkingmachinesInkling-Small is a lighter-weight model with 12B active parameters, trained with a similar recipe, to Inkling that achieves strong performance with even lower cost and latency.Gemini 3.6 FlashGoogleGemini 3.6 Flash delivers higher quality across coding, agentic workflows, and web development with reduced token consumption and fewer model calls compared to previous model iterations.InklingThinkingmachinesInkling is a multimodal MoE model (975B total, 41B active, 256k context) reasoning over text, image, and audio inputs.
Available with one subscription
Start with Gemini 3.5 Flash Lite in Shortly AI
Keep your conversations, files and model choices together in one workspace.