Google Launches Gemini 3.6 Flash with 17% Output Token Efficiency Gain
Google debuts three new Gemini models including 3.6 Flash with improved efficiency, while 3.5 Pro remains in limited testing.
Gemini 3.6 Flash: Efficiency and Performance Gains
Google has released Gemini 3.6 Flash, delivering a 17% reduction in output token usage compared to Gemini 3.5 Flash according to the Artificial Analysis Index. The model is priced at $1.50 per million input tokens and $7.50 per million output tokens.
Across multiple benchmarks, Gemini 3.6 Flash demonstrates meaningful improvements:
- DeepSWE: 49% versus 37% on Gemini 3.5 Flash
- OSWorld-Verified: 83.0% versus 78.4% on Gemini 3.5 Flash
- GDPval-AA v2: 1421 versus 1349 on Gemini 3.5 Flash
- MLE Bench: 63.9% versus 49.7% on Gemini 3.5 Flash
Gemini 3.5 Flash-Lite: Speed and Cost Optimized
Google has also introduced Gemini 3.5 Flash-Lite, a lightweight model delivering 350 output tokens per second according to the Artificial Analysis Index. It is priced at $0.3 per million input tokens and $2.5 per million output tokens.
The model shows substantial gains over Gemini 3.1 Flash-Lite:
- Terminal-Bench 2.1: 54% versus 31%
- GDM-MRCR v2: 72.2% versus 60.1%
- GDPval-AA v2: 1140 versus 642
Compared to Gemini 3 Flash, Gemini 3.5 Flash-Lite achieves:
- SWE-Bench Pro: 54.2% versus 49.6%
- OSWorld-Verified: 74.0% versus 65.1%
Specialized and Forthcoming Models
Gemini 3.5 Flash Cyber will be exclusively available to governments and trusted partners as part of a limited-access pilot program.
Gemini 3.5 Pro is currently testing with partners, with Google planning to make it broadly available as soon as it’s ready.
Google has started its most ambitious pre-training run yet for Gemini 4.
Source: Google Official Blog
Irish pronunciation
All FoxxeLabs components are named in Irish. Click ▶ to hear each name spoken by a native Irish voice.