Gemini 3.6 Flash
Latency-sensitive coding and agentic tasks that still need a huge context window.
Gemini 3.6 Flash supersedes Gemini 3.5 Flash, Gemini 3.1 Flash, Gemini 2.5 Flash.
What are Gemini 3.6 Flash's specs and price?
Gemini 3.6 Flash, built by Google, ships a 1M-token context window and a 64K-token max output, released 2026-06. It supports text and vision and audio input with a dedicated reasoning mode and costs $3.38 per million blended tokens, the 26th-cheapest of 32 models we track.
Specs
| Context window | 1M tokens |
| Max output | 64K tokens |
| Modalities | text, vision, audio |
| Extended thinking | Yes |
| Released | 2026-06 |
| Knowledge cutoff | 2026-04 |
| Provider |
Verified 2026-08-08 — source.
Where it ranks
Strengths
- Faster than Gemini 3.5 Flash
- Top-tier coding and agentic performance for its price
- 1M-token context
More about Gemini 3.6 Flash
FAQ
What is Gemini 3.6 Flash's context window?
Gemini 3.6 Flash has a 1M-token context window and a 64K-token max output — the 6th-largest context of the 32 current models we track. Source: https://ai.google.dev/gemini-api/docs/models, verified 2026-08-08.
Does Gemini 3.6 Flash support vision or audio input?
Yes — Gemini 3.6 Flash accepts vision and audio input in addition to text.
Does Gemini 3.6 Flash have a reasoning or extended-thinking mode?
Yes — Gemini 3.6 Flash exposes a dedicated reasoning mode for multi-step problems.
When was Gemini 3.6 Flash released, and what is its knowledge cutoff?
Gemini 3.6 Flash was released 2026-06 with a knowledge cutoff of 2026-04.
How much does Gemini 3.6 Flash cost, and who provides it?
Gemini 3.6 Flash is served by Google at $3.38/M blended tokens (3:1 input:output) — the 26th-cheapest of 32 current models. Full pricing breakdown: /llm-api-pricing/gemini-3-6-flash.
