Skip to content

Gemini 3.7 Flash

gemini-3.7-flash
Google · chat · api
GA Alert me on changes
Intelligence
Context
1M
Max output
65.5K
Input $/1M
$0.75
Output $/1M
$3.75
Modalities
text
Released
13 Aug 2026
Intelligence Index via Artificial Analysis · 0–100, higher is better
Download image Share on X Share on LinkedIn
AI summary
● machine-written

Google releases Gemini 3.7 Flash for agentic workflows and coding

Gemini 3.7 Flash is a multimodal model from Google supporting text, image, video, audio, and PDF inputs with a 1M token context window. It is designed for fast agentic workflows, coding, and complex multi-step reasoning tasks. The model is available through Google's API with introductory pricing through December 31, 2026.

What's new
  • 1M token context window with 65,536 token maximum output
  • Supports function calling, structured outputs, and thinking modes
  • Introductory pricing of $0.75/$3.75 per 1M input/output tokens through end of 2026
  • Supports caching, code execution, computer use, and file search
  • Released August 13, 2026
Best for
Agentic workflows requiring responsive performanceComplex coding and multi-step reasoning tasksFast, cost-effective inference with large context requirementsMulti-modal input processing including video and audio
Sources

Common questions about Gemini Flash →

Source: https://ai.google.dev/gemini-api/docs/models/gemini-3.7-flash