Gemini 3.8 Flash
Newest and recommended Gemini text model as of 2026-09 (no Pro newer than 3.1 Pro preview; 3.5 Pro announced but unreleased). Aliases gemini-flash-latest may point here. Model card says some domains' knowledge only to 2025-01. Vertex id inferred from docs page.
- Context window
- 1,048,576 tokens
- Max output
- 65,536 tokens
- Knowledge cutoff
- 2026-03
- Input
- text, image, audio, video, pdf
- Output
- text
- License
- proprietary
- Pricing
- input: $0.75 · output: $3.75 (per 1M tokens (Standard tier, introductory price through 2026-12-31; rises to $1.50 / $7.50 from 2027-01-01)) source
- Verified
- 2026-09-29
How to call it
| Provider | Model id | Endpoint / URL | Docs |
|---|---|---|---|
| Gemini API | gemini-3.8-flash | https://generativelanguage.googleapis.com/v1beta/models/gemini-3.8-flash:generateContent | docs |
| Google Cloud Gemini Enterprise Agent Platform (Vertex AI) | gemini-3.8-flash | — | docs |
| OpenRouter | google/gemini-3.8-flash | — | — |
| Web app | — | gemini.google.com | — |
Notable capabilities (5)
- Long-horizon software engineering: Google's most capable Flash for autonomous end-to-end engineering; 73.7% on DeepSWE v1.1, 89.4% terminal-based coding per model card. source
- Specialized-domain agentic analysis: Beats 3.7 Flash and other frontier models on Vals Finance Agent v2 (61.4%) and Harvey's Legal Agent Benchmark. source
- Agentic long-video understanding: 87.8% long video understanding in agentic mode (agentic video understanding added for 3.x Flash on 2026-09-01). source
- Adjustable thinking levels + computer use: Thinking low/medium/high, computer use (preview), Maps/Search grounding, flex and priority inference tiers. source
- Cyber sibling model: Launched alongside Gemini 3.8 Flash Cyber (vulnerability detection/patching, 47.2% CWE-Bench pass@1), available only to vetted defenders via the Fairwind Program. source
Google's flagship workhorse model (Sept 2026): best Gemini for coding agents, multi-step reasoning and long multimodal context at Flash prices.
curl "https://generativelanguage.googleapis.com/v1beta/models/gemini-3.8-flash:generateContent" \
-H "x-goog-api-key: $GEMINI_API_KEY" -H "Content-Type: application/json" \
-d '{"contents":[{"parts":[{"text":"Explain how AI works in a few words"}]}]}'
Sources: model page, pricing, launch blog, model card.
Other Google DeepMind models
Gemini 3.8 Flash TTS · Gemini 3.8 Live · Gemini 3.5 Transcribe (and Transcribe Live) · Lyria 3.5 · Gemini 3.5 Flash-Lite · Nano Banana 2 Lite (Gemini 3.1 Flash-Lite Image) · Gemini Omni Flash (Omni 1.1 Flash) · Gemini Embedding 2 · Gemma 4 · Nano Banana 2 (Gemini 3.1 Flash Image) · Nano Banana Pro (Gemini 3 Pro Image) · Gemini Robotics 2 · Gemini Robotics ER 2 · Gemini Robotics On-Device 2 · Gemini 3.5 Live Translate · Gemini 3.1 Pro · Veo 3.1 · Genie 3 · Lyria RealTime · Gemini 3.7 Flash · Gemini 3.6 Flash · Gemini 3.5 Flash · Gemini 3.1 Flash TTS (preview) · Gemini 3.1 Flash Live (preview) · Lyria 3 (Clip / Pro) · Lyria 2 · Gemini 2.5 Flash Native Audio (Live, preview) · Gemini 2.5 Flash-Lite · Gemini 2.5 Flash · Gemini 2.5 Pro · Gemini 2.5 Flash TTS / Pro TTS · Gemini 3.1 Flash-Lite · Nano Banana (Gemini 2.5 Flash Image) · Gemini Robotics-ER 1.5 / 1.6 · Imagen 4