Gemini 3.6 Flash is optimized for multi-step orchestration, full-stack code refactoring, and general reasoning. It improves on many key areas critical to the Flash line of models.
Improvements from previous Flash models include:
- Improved token efficiency over Gemini 3.5 Flash. 3.6 Flash uses less tokens and completes multi-step workflows in fewer turns.
- Improved code generation with lower compile-failure and revision rates across building, prototyping, and IDE agent environments.
- Reduces action bias by resolving read-only diagnostic tasks without making unsolicited edits.
- Improved multimodal reasoning with respect to chart interpretation, visual blueprint conversion, and multi-element web layout generation.
Flash includes the following potentially breaking changes when compared to previous Gemini models:
- Custom values for parameters like temperature, top-K, and top-P aren't supported. If you set a custom value for these parameters, that value will be ignored.
- Custom values for frequency and presence penalty parameters aren't supported. Setting a custom value for these parameters will throw an error.
- API requests where the last input turn has a role of
Modelaren't supported. The following kinds of requests will return an error:- When using the Interactions API: Requests where the last object in the
inputarray has"type": "model_output". - When using the GenerateContent API: Requests where the last object in
the
contentsarray has"role": "model".
- When using the Interactions API: Requests where the last object in the
Try in Agent Studio Deploy example app Developer guide Pricing
| Model ID | gemini-3.6-flash |
|
|---|---|---|
| Modalities |
|
|
| Token limits | Context window | 1,048,576 |
| Maximum output tokens | 65,536 | |
| Capabilities |
|
|
| Tools |
|
|
| APIs |
|
|
| Consumption options |
|
|
| Technical specifications | Image |
|
| Text |
|
|
| Video |
|
|
| Audio |
|
|
| Parameter defaults |
|
|
| Supported regions |
|
|
|
||
|
||
|
||
| Versions |
|
|
| Security controls | Online prediction |
|
| Batch inference |
|
|
| Context caching |
|
|
| See Security controls for more information. | ||
† Listed retirement dates refer to retirement of support in Gemini Enterprise Agent Platform. Models may remain accessible through the Gemini API after these dates have passed. The Gemini API is not a Google Cloud offering and is subject to its own terms of service. For details, see the Gemini API documentation.