Gemini 3.8 Flash
Autonomous agents and enterprise coding workflows that need Flash speed without Pro pricing.
Gemini 3.8 Flash supersedes Gemini 3.6 Flash, Gemini 3.5 Flash, Gemini 3.1 Flash, Gemini 2.5 Flash.
What are Gemini 3.8 Flash's specs and price?
Gemini 3.8 Flash, built by Google, ships a 1.0M-token context window and a 66K-token max output, released 2026-09. It supports text and vision and audio input with a dedicated reasoning mode and costs $1.50 per million blended tokens, the 22nd-cheapest of 49 models we track.
What are Gemini 3.8 Flash's specs?
| Context window | 1.0M tokens |
| Max output | 66K tokens |
| Modalities | text, vision, audio |
| Extended thinking | Yes |
| Released | 2026-09 |
| Knowledge cutoff | Not published |
| Provider |
Verified 2026-10-01 — source.
Where does Gemini 3.8 Flash rank?
What are Gemini 3.8 Flash's strengths?
- Most intelligent Flash model to date
- Long-horizon software engineering at Flash cost
- 1M-token context with built-in tools
What else should you know about Gemini 3.8 Flash?
What are common questions about Gemini 3.8 Flash?
What is Gemini 3.8 Flash's context window?
Gemini 3.8 Flash has a 1.0M-token context window and a 66K-token max output — the 8th-largest context of the 49 current models we track. Source: https://ai.google.dev/gemini-api/docs/models/gemini-3.8-flash, verified 2026-10-01.
Does Gemini 3.8 Flash support vision or audio input?
Yes — Gemini 3.8 Flash accepts vision and audio input in addition to text.
Does Gemini 3.8 Flash have a reasoning or extended-thinking mode?
Yes — Gemini 3.8 Flash exposes a dedicated reasoning mode for multi-step problems.
When was Gemini 3.8 Flash released, and what is its knowledge cutoff?
Gemini 3.8 Flash was released 2026-09.
How much does Gemini 3.8 Flash cost, and who provides it?
Gemini 3.8 Flash is served by Google at $1.50/M blended tokens (3:1 input:output) — the 22nd-cheapest of 49 current models. Full pricing breakdown: /llm-api-pricing/gemini-3-8-flash.
