Skip to content

Gemini 3.8 Flash

gemini-3.8-flash
Google · chat · api
GA Alert me on changes
Intelligence
Context
1M
Max output
65.5K
Input $/1M
$0.75
Output $/1M
$3.75
Modalities
text
Released
02 Sept 2026
Intelligence Index via Artificial Analysis · 0–100, higher is better
Download image Share on X Share on LinkedIn
AI summary
● machine-written

Google releases Gemini 3.8 Flash for long-context tasks

Gemini 3.8 Flash is Google's most intelligent Flash model, designed for long-horizon software engineering, autonomous agents, and complex enterprise workflows. It supports a 1-million-token context window with 65,536 token maximum output. The model is available via API with introductory pricing through December 31, 2026.

What's new
  • 1M-token context window for long-horizon tasks
  • Supports computer use (Preview) and structured outputs
  • Introductory pricing of $0.75/$3.75 per 1M tokens through Dec 31, 2026
  • Available through Gemini API, Google AI Studio, and developer platforms
  • Supports Batch API, Flex inference, and Priority inference
Best for
Software engineering and code generationAutonomous agents and agentic workflowsComplex multi-step enterprise reasoningLong-context document analysis
Sources

Common questions about Gemini Flash →

Source: https://ai.google.dev/gemini-api/docs/models/gemini-3.8-flash