Gemini 3.6 Flash (GA)
Gemini 3.6 Flash is a Google model that became generally available in the Gemini API on July 21, 2026.
- Lab
- Google DeepMind
- Release date
- Status
- Generally available
- API model ID
- gemini-3.6-flash
- Context window
- 1,048,576 tokens
- Max output
- 65,536 tokens
Google made Gemini 3.6 Flash generally available in the Gemini API on July 21, 2026, citing improved token efficiency. The Gemini API docs list it as gemini-3.6-flash, with text, image, video, audio and PDF input, text output, a 1,048,576-token input limit and a 65,536-token output limit.
- Modalities
- Text, image, video, audio and PDF input; text output
- Availability
- Gemini API.
- Official source
- ai.google.dev
- Last checked
Questions
When did Gemini 3.6 Flash become generally available?
Google made Gemini 3.6 Flash generally available in the Gemini API on July 21, 2026.
What is the context window of Gemini 3.6 Flash?
The Gemini API docs list a 1,048,576-token input limit and a 65,536-token output limit.
What changed in Gemini 3.6 Flash?
Google cited improved token efficiency when it made the model generally available.
Source
Release dates, model IDs and limits come from each lab’s own announcement, changelog, documentation or deprecations page, linked on every model with the date it was last checked. How we pick and check sources: editorial standards.