Gemini 3.5 Flash Lite
High-volume tasks that still need long-context handling on a tight budget.
Gemini 3.5 Flash Lite supersedes Gemini 3.1 Flash Lite, Gemini 2.5 Flash Lite.
What are Gemini 3.5 Flash Lite's specs and price?
Gemini 3.5 Flash Lite, built by Google, ships a 1M-token context window and a 32K-token max output, released 2026-04. It supports text and vision input and costs $0.56 per million blended tokens, the 12th-cheapest of 32 models we track.
Specs
| Context window | 1M tokens |
| Max output | 32K tokens |
| Modalities | text, vision |
| Extended thinking | No |
| Released | 2026-04 |
| Knowledge cutoff | 2026-02 |
| Provider |
Verified 2026-08-08 — source.
Where it ranks
Strengths
- Cheapest Gemini with a 1M-token context window
- Faster and cheaper than 3.1 Flash Lite
- Vision input included
More about Gemini 3.5 Flash Lite
FAQ
What is Gemini 3.5 Flash Lite's context window?
Gemini 3.5 Flash Lite has a 1M-token context window and a 32K-token max output — the 7th-largest context of the 32 current models we track. Source: https://ai.google.dev/gemini-api/docs/models, verified 2026-08-08.
Does Gemini 3.5 Flash Lite support vision or audio input?
Yes — Gemini 3.5 Flash Lite accepts vision input in addition to text.
Does Gemini 3.5 Flash Lite have a reasoning or extended-thinking mode?
No — Gemini 3.5 Flash Lite does not expose a separate reasoning/extended-thinking mode.
When was Gemini 3.5 Flash Lite released, and what is its knowledge cutoff?
Gemini 3.5 Flash Lite was released 2026-04 with a knowledge cutoff of 2026-02.
How much does Gemini 3.5 Flash Lite cost, and who provides it?
Gemini 3.5 Flash Lite is served by Google at $0.56/M blended tokens (3:1 input:output) — the 12th-cheapest of 32 current models. Full pricing breakdown: /llm-api-pricing/gemini-3-5-flash-lite.
