Gemini 3.5 Flash-Lite (GA)
Gemini 3.5 Flash-Lite is a Google model that became generally available in the Gemini API on July 21, 2026.
- Lab
- Google DeepMind
- Release date
- Status
- Generally available
- API model ID
- gemini-3.5-flash-lite
- Context window
- 1,048,576 tokens
- Max output
- 65,536 tokens
Google made Gemini 3.5 Flash-Lite generally available in the Gemini API on July 21, 2026, the same day as Gemini 3.6 Flash. The Gemini API docs list it as gemini-3.5-flash-lite, with text, image, video, audio and PDF input, text output, a 1,048,576-token input limit and a 65,536-token output limit.
- Modalities
- Text, image, video, audio and PDF input; text output
- Availability
- Gemini API.
- Official source
- ai.google.dev
- Last checked
Questions
When did Gemini 3.5 Flash-Lite become generally available?
Google made it generally available in the Gemini API on July 21, 2026, the same day as Gemini 3.6 Flash.
What is the context window of Gemini 3.5 Flash-Lite?
The Gemini API docs list a 1,048,576-token input limit and a 65,536-token output limit.
What inputs does Gemini 3.5 Flash-Lite accept?
It accepts text, image, video, audio and PDF input and returns text output.
Source
Release dates, model IDs and limits come from each lab’s own announcement, changelog, documentation or deprecations page, linked on every model with the date it was last checked. How we pick and check sources: editorial standards.