Gemini 3.7 Flash
Production coding and agentic workflows that need strong capability, low latency, and long context.
What are Gemini 3.7 Flash's specs and price?
Gemini 3.7 Flash, built by Google, ships a 1.0M-token context window and a 66K-token max output, released 2026-08. It supports text and vision and audio input with a dedicated reasoning mode and costs $1.50 per million blended tokens, the 15th-cheapest of 37 models we track.
gemini-3.7-flash · Read the release analysis →Specs
| Context window | 1.0M tokens |
| Max output | 66K tokens |
| Modalities | text, vision, audio |
| Extended thinking | Yes |
| Released | 2026-08 |
| Knowledge cutoff | Not published |
| Provider | |
| Tools | function calling, code execution, search grounding, file search, structured output, computer use (preview) |
Verified 2026-08-14 — source.
Where it ranks
Strengths
- Google’s most intelligent Flash workhorse
- Tunable thinking for coding and agents
- 1M-token context with built-in tools
More about Gemini 3.7 Flash
FAQ
What is Gemini 3.7 Flash's context window?
Gemini 3.7 Flash has a 1.0M-token context window and a 66K-token max output — the 2nd-largest context of the 37 current models we track. Source: https://ai.google.dev/gemini-api/docs/models/gemini-3.7-flash, verified 2026-08-14.
Does Gemini 3.7 Flash support vision or audio input?
Yes — Gemini 3.7 Flash accepts vision and audio input in addition to text.
Does Gemini 3.7 Flash have a reasoning or extended-thinking mode?
Yes — Gemini 3.7 Flash exposes a dedicated reasoning mode for multi-step problems.
When was Gemini 3.7 Flash released, and what is its knowledge cutoff?
Gemini 3.7 Flash was released 2026-08.
How much does Gemini 3.7 Flash cost, and who provides it?
Gemini 3.7 Flash is served by Google at $1.50/M blended tokens (3:1 input:output) — the 15th-cheapest of 37 current models. Full pricing breakdown: /llm-api-pricing/gemini-3-7-flash.
