Builders and prospects constructing manufacturing AI brokers want larger token effectivity, decrease latency, and extra dependable efficiency. Our Flash sequence of fashions is constructed to fulfill the candy spot of effectivity and high quality to allow scaling agentic workflows. Constructing on Gemini 3.5 Flash, we’re introducing new Gemini fashions:
- 3.6 Flash: Our workhorse mannequin that delivers higher coding, information work, and multimodal efficiency. In keeping with the Synthetic Evaluation Index, it reduces output token utilization by 17% in comparison with 3.5 Flash, and in some benchmarks like DeepSWE by Datacurve, we observe as much as 65%, all at a decrease value per output token.
- 3.5 Flash-Lite: Our quickest, most cost-effective 3.5-class mannequin, delivering 350 output tokens per second in line with the Synthetic Evaluation Index, additionally considerably outperforming prior Flash-Lite generations in agentic workflows.
- 3.5 Flash Cyber in CodeMender: Profitable cybersecurity functions require cautious orchestration of a mannequin alongside an agent infrastructure. We’re introducing a mix of a brand new, extremely environment friendly, specialised cyber-focused mannequin paired with our CodeMender code safety agent that delivers aggressive efficiency on the frontier.
Past immediately’s releases, Gemini 3.5 Professional is at the moment testing with companions and we plan to make it broadly out there as quickly because it’s prepared. In parallel, our group is already specializing in constructing the subsequent era of fashions. We’ve got began our most bold pre-training run but, for Gemini 4, and are excited by the progress.
3.6 Flash: Extra environment friendly and higher high quality than 3.5 Flash
Gemini 3.6 Flash builds instantly on developer and buyer suggestions from 3.5 Flash. 3.6 Flash not solely delivers a step up in coding and information work, however it does this whereas meaningfully enhancing token effectivity. For instance, on the Synthetic Evaluation Index, we see 3.6 Flash consuming 17% fewer output tokens than 3.5 Flash. It additionally takes fewer reasoning steps and power calls to perform multi-step workflows.
This enhanced effectivity can also be mixed with a lower cost than 3.5 Flash. At $1.50/1M enter tokens and $7.50/1M output tokens, 3.6 Flash reduces the general value per agentic job, making brokers more cost effective to construct and run.









