Gemini 3.6 Flash: Efficiency and Performance Gains

Google has released Gemini 3.6 Flash, delivering a 17% reduction in output token usage compared to Gemini 3.5 Flash according to the Artificial Analysis Index. The model is priced at $1.50 per million input tokens and $7.50 per million output tokens.

Across multiple benchmarks, Gemini 3.6 Flash demonstrates meaningful improvements:

  • DeepSWE: 49% versus 37% on Gemini 3.5 Flash
  • OSWorld-Verified: 83.0% versus 78.4% on Gemini 3.5 Flash
  • GDPval-AA v2: 1421 versus 1349 on Gemini 3.5 Flash
  • MLE Bench: 63.9% versus 49.7% on Gemini 3.5 Flash

Gemini 3.5 Flash-Lite: Speed and Cost Optimized

Google has also introduced Gemini 3.5 Flash-Lite, a lightweight model delivering 350 output tokens per second according to the Artificial Analysis Index. It is priced at $0.3 per million input tokens and $2.5 per million output tokens.

The model shows substantial gains over Gemini 3.1 Flash-Lite:

  • Terminal-Bench 2.1: 54% versus 31%
  • GDM-MRCR v2: 72.2% versus 60.1%
  • GDPval-AA v2: 1140 versus 642

Compared to Gemini 3 Flash, Gemini 3.5 Flash-Lite achieves:

  • SWE-Bench Pro: 54.2% versus 49.6%
  • OSWorld-Verified: 74.0% versus 65.1%

Specialized and Forthcoming Models

Gemini 3.5 Flash Cyber will be exclusively available to governments and trusted partners as part of a limited-access pilot program.

Gemini 3.5 Pro is currently testing with partners, with Google planning to make it broadly available as soon as it’s ready.

Google has started its most ambitious pre-training run yet for Gemini 4.


Source: Google Official Blog