Model · Google · Text
Gemini 3.8 Flash
Google's most intelligent Flash model — engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows at Flash-tier speed and cost. The current head of the Gemini Flash line.
- Modality
- Text
- License
- Proprietary (Proprietary)
- Context window
- 1,048,576 tokens 1M-token input with 64k max output. Thinking levels are low / medium (default) / high — the minimal thinking level is not supported.
- Released
- September 2, 2026
- Last verified
- September 8, 2026
- Runs locally
- No
- Also handles
- vision, reasoning, code
Strengths
- Google's best reasoning and coding model at the same speed and low cost as Gemini 3.7 Flash
- 1M-token context across text, images, video, audio, and documents
- Built for long-horizon software engineering, autonomous agents, and complex enterprise workflows
- Introductory pricing of $0.75 / $3.75 per MTok through 2026-12-31
Weaknesses
- Uses more tokens on longer, more complex tasks by design — Google points everyday work at a lower reasoning effort or Gemini 3.7 Flash
- The minimal thinking level is not supported
- List pricing rises to $1.50 / $7.50 per MTok on 2027-01-01
- Output capped at 64k tokens
- Closed weights
Try it
| Where | Type | Notes |
|---|---|---|
| Google AI Studio | hosted-api | Free tier |
| Vertex AI | hosted-api | GCP enterprise |
Used in solutions
Version history
- Gemini 3.8 Flash Sep 2026 Current
- Gemini 3.5 Flash-Lite Jul 2026
- Gemini 3.5 Flash May 2026 Deprecated
- Gemini 3.1 Flash-Lite May 2026 Deprecated
- Gemini 3.1 Pro Preview May 2026
Official sources
- Model docs docs
- Announcement announcement
- Model card model
Change log
- — Initial entry. Released 2026-09-02 per Google's Gemini API changelog as the head of the Flash line, superseding Gemini 3.7 Flash (2026-08-13) and 3.6 Flash (2026-07-21). Dates taken from the changelog and deprecations pages, not the models page (whose per-model stamps are a global page timestamp).
Esc