Gemini 3.6 Flash
gemini-3.6-flash
Google · chat · api
Intelligence
Context
1M
Max output
65.5K
Input $/1M
$1.50
Output $/1M
$7.50
Modalities
text
Released
21 Jul 2026
Intelligence Index via Artificial Analysis · 0–100, higher is better
AI summary
● machine-written
Google releases Gemini 3.6 Flash for coding and agentic workflows
Gemini 3.6 Flash is Google's latest model in the Flash series, designed for the agentic era with focus on coding, multi-step workflows, and real-world tasks. It supports a 1M token context window with 65.5K maximum output and is available via API at specified pricing. The model emphasizes efficiency through reduced token use and fewer model calls needed to complete tasks.
What's new
- Released July 21, 2026
- 1M token context window with 65.5K output limit
- Supports prompt caching to reduce effective input cost by up to 90%
- Excels at coding, agentic workflows, and spatial reasoning
- Designed for fast agent loops with complex coding cycles
Best for
Code generation and software developmentAgentic workflows and multi-step reasoningWeb and app developmentSpatial reasoning tasks
Common questions about Gemini Flash →
Source: https://ai.google.dev/gemini-api/docs/models/gemini-3.6-flash