Gemini 3.5 Flash Lite
High-efficiency multimodal model for focused subagent tasks.
Gemini 3.5 Flash Lite is a high-efficiency model with upgraded agentic capabilities, suited for subagents that execute focused tasks within complex multi-agent workflows.
- Provider
- Type
- Language model
- Released
- Jul 21, 2026
- Context window
- 1M tokens
- Max output
- 66K tokens
- Price
- $0.30 input / $2.50 output per 1M tokens
- Input
- text, image, audio, video, file
- Output
- text
- Open weights
- No
- Knowledge cutoff
- Mar 31, 2026
Best for
- High volume / low cost
- Multimodal
- Extraction & classification
- Agents
Strengths
- Low price with full multimodality
- 1M context
More from Google
- Gemini 4 Argon — Google. Google’s next-generation frontier model, currently restricted to vetted cyber defenders.
- Gemini 3.8 Flash — Google. Google’s most intelligent Flash model — fully multimodal, fast and affordable.
- Nano Banana 2.1 — Google. Google’s latest image generation and editing model.
- Gemma 4 31B — Google. Open-weight dense multimodal model from Google DeepMind.
- Gemma 4 26B A4B — Google. Open MoE Gemma: ~31B-class quality with only 3.8B active parameters.
- Gemini Embedding 2 — Google. Google’s first multimodal embedding model.
- Gemini Omni Flash 1.1 — Google. Google’s fast, low-cost video generation model.
- Veo 3.1 — Google. Google’s Veo video model with native audio; powers Flow.