Gemini 3.7 Flash
gemini-3.7-flash
Google · chat · api
Intelligence
Context
1M
Max output
65.5K
Input $/1M
$0.75
Output $/1M
$3.75
Modalities
text
Released
13 Aug 2026
Intelligence Index via Artificial Analysis · 0–100, higher is better
AI summary
● machine-written
Google releases Gemini 3.7 Flash for agentic workflows and coding
Gemini 3.7 Flash is a multimodal model from Google supporting text, image, video, audio, and PDF inputs with a 1M token context window. It is designed for fast agentic workflows, coding, and complex multi-step reasoning tasks. The model is available through Google's API with introductory pricing through December 31, 2026.
What's new
- 1M token context window with 65,536 token maximum output
- Supports function calling, structured outputs, and thinking modes
- Introductory pricing of $0.75/$3.75 per 1M input/output tokens through end of 2026
- Supports caching, code execution, computer use, and file search
- Released August 13, 2026
Best for
Agentic workflows requiring responsive performanceComplex coding and multi-step reasoning tasksFast, cost-effective inference with large context requirementsMulti-modal input processing including video and audio
Common questions about Gemini Flash →
Source: https://ai.google.dev/gemini-api/docs/models/gemini-3.7-flash