Gemini 3.8 Flash
gemini-3.8-flash
Google · chat · api
Intelligence
Context
1M
Max output
65.5K
Input $/1M
$0.75
Output $/1M
$3.75
Modalities
text
Released
02 Sept 2026
Intelligence Index via Artificial Analysis · 0–100, higher is better
AI summary
● machine-written
Google releases Gemini 3.8 Flash for long-context tasks
Gemini 3.8 Flash is Google's most intelligent Flash model, designed for long-horizon software engineering, autonomous agents, and complex enterprise workflows. It supports a 1-million-token context window with 65,536 token maximum output. The model is available via API with introductory pricing through December 31, 2026.
What's new
- 1M-token context window for long-horizon tasks
- Supports computer use (Preview) and structured outputs
- Introductory pricing of $0.75/$3.75 per 1M tokens through Dec 31, 2026
- Available through Gemini API, Google AI Studio, and developer platforms
- Supports Batch API, Flex inference, and Priority inference
Best for
Software engineering and code generationAutonomous agents and agentic workflowsComplex multi-step enterprise reasoningLong-context document analysis
Common questions about Gemini Flash →
Source: https://ai.google.dev/gemini-api/docs/models/gemini-3.8-flash