Gemini 3.6 Flash (GA)

Gemini 3.6 Flash is a Google model that became generally available in the Gemini API on July 21, 2026.

Lab
Google DeepMind
Release date
Status
Generally available
API model ID
gemini-3.6-flash
Context window
1,048,576 tokens
Max output
65,536 tokens

Google made Gemini 3.6 Flash generally available in the Gemini API on July 21, 2026, citing improved token efficiency. The Gemini API docs list it as gemini-3.6-flash, with text, image, video, audio and PDF input, text output, a 1,048,576-token input limit and a 65,536-token output limit.

Modalities
Text, image, video, audio and PDF input; text output
Availability
Gemini API.
Official source
ai.google.dev
Last checked

Questions

When did Gemini 3.6 Flash become generally available?
Google made Gemini 3.6 Flash generally available in the Gemini API on July 21, 2026.
What is the context window of Gemini 3.6 Flash?
The Gemini API docs list a 1,048,576-token input limit and a 65,536-token output limit.
What changed in Gemini 3.6 Flash?
Google cited improved token efficiency when it made the model generally available.

Source

Release dates, model IDs and limits come from each lab’s own announcement, changelog, documentation or deprecations page, linked on every model with the date it was last checked. How we pick and check sources: editorial standards.

The Model Press

What are you looking for?

Search by headline, topic or keyword.