Skip to content

Gemini 3.6 Flash

gemini-3.6-flash
Google · chat · api
GA Alert me on changes
Intelligence
Context
1M
Max output
65.5K
Input $/1M
$1.50
Output $/1M
$7.50
Modalities
text
Released
21 Jul 2026
Intelligence Index via Artificial Analysis · 0–100, higher is better
Download image Share on X Share on LinkedIn
AI summary
● machine-written

Google releases Gemini 3.6 Flash for coding and agentic workflows

Gemini 3.6 Flash is Google's latest model in the Flash series, designed for the agentic era with focus on coding, multi-step workflows, and real-world tasks. It supports a 1M token context window with 65.5K maximum output and is available via API at specified pricing. The model emphasizes efficiency through reduced token use and fewer model calls needed to complete tasks.

What's new
  • Released July 21, 2026
  • 1M token context window with 65.5K output limit
  • Supports prompt caching to reduce effective input cost by up to 90%
  • Excels at coding, agentic workflows, and spatial reasoning
  • Designed for fast agent loops with complex coding cycles
Best for
Code generation and software developmentAgentic workflows and multi-step reasoningWeb and app developmentSpatial reasoning tasks
Sources

Common questions about Gemini Flash →

Source: https://ai.google.dev/gemini-api/docs/models/gemini-3.6-flash