Cyberax AI Playbook
cyberax.com
Model · Google · Text

Gemini 3.8 Flash

Google's most intelligent Flash model — engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows at Flash-tier speed and cost. The current head of the Gemini Flash line.

Modality
Text
License
Proprietary (Proprietary)
Context window
1,048,576 tokens 1M-token input with 64k max output. Thinking levels are low / medium (default) / high — the minimal thinking level is not supported.
Released
September 2, 2026
Last verified
September 8, 2026
Runs locally
No
Also handles
vision, reasoning, code

Strengths

  • Google's best reasoning and coding model at the same speed and low cost as Gemini 3.7 Flash
  • 1M-token context across text, images, video, audio, and documents
  • Built for long-horizon software engineering, autonomous agents, and complex enterprise workflows
  • Introductory pricing of $0.75 / $3.75 per MTok through 2026-12-31

Weaknesses

  • Uses more tokens on longer, more complex tasks by design — Google points everyday work at a lower reasoning effort or Gemini 3.7 Flash
  • The minimal thinking level is not supported
  • List pricing rises to $1.50 / $7.50 per MTok on 2027-01-01
  • Output capped at 64k tokens
  • Closed weights

Try it

WhereTypeNotes
Google AI Studio hosted-api Free tier
Vertex AI hosted-api GCP enterprise

Used in solutions

Version history

  1. Gemini 3.8 Flash Sep 2026 Current
  2. Gemini 3.5 Flash-Lite Jul 2026
  3. Gemini 3.5 Flash May 2026 Deprecated
  4. Gemini 3.1 Flash-Lite May 2026 Deprecated
  5. Gemini 3.1 Pro Preview May 2026

Change log

  • — Initial entry. Released 2026-09-02 per Google's Gemini API changelog as the head of the Flash line, superseding Gemini 3.7 Flash (2026-08-13) and 3.6 Flash (2026-07-21). Dates taken from the changelog and deprecations pages, not the models page (whose per-model stamps are a global page timestamp).