Gemini 3.8 Flash
Google’s most intelligent Flash model — fully multimodal, fast and affordable.
Gemini 3.8 Flash is Google’s most intelligent Flash model, with significant gains over 3.7 Flash across software engineering, agentic tasks and multi-step reasoning. It accepts text, image, audio, video and files.
- Provider
- Type
- Language model
- Released
- Sep 2, 2026
- Context window
- 1M tokens
- Max output
- 66K tokens
- Price
- $0.75 input / $3.75 output per 1M tokens
- Input
- text, image, audio, video, file
- Output
- text
- Open weights
- No
- Knowledge cutoff
- Mar 31, 2026
Best for
- Multimodal
- Vision
- Long context
- Real-time / low latency
- Agents
- Math & science
Strengths
- Native audio & video input
- Very high GPQA for a Flash model
- Fast
More from Google
- Gemini 4 Argon — Google. Google’s next-generation frontier model, currently restricted to vetted cyber defenders.
- Gemini 3.5 Flash Lite — Google. High-efficiency multimodal model for focused subagent tasks.
- Nano Banana 2.1 — Google. Google’s latest image generation and editing model.
- Gemma 4 31B — Google. Open-weight dense multimodal model from Google DeepMind.
- Gemma 4 26B A4B — Google. Open MoE Gemma: ~31B-class quality with only 3.8B active parameters.
- Gemini Embedding 2 — Google. Google’s first multimodal embedding model.
- Gemini Omni Flash 1.1 — Google. Google’s fast, low-cost video generation model.
- Veo 3.1 — Google. Google’s Veo video model with native audio; powers Flow.
Tools that use it
- GitHub Copilot — GitHub (Microsoft). The most widely deployed AI pair programmer.
- Gemini CLI — Google. Open-source terminal coding agent powered by Gemini.
- Agent Development Kit — Google. Google’s open-source framework for multi-agent systems.
- Vertex AI — Google Cloud. Google Cloud’s AI platform.
- Gemini — Google. Google’s AI assistant across Search, Android and Workspace.