This model is 19 months old. A newer version of it has since shipped: Gemini 3.8 Flash (2026-09-02). The details below are still accurate for this model — it just is not what we would recommend today.

Gemini 2.0 Flash-Lite

Google DeepMind🇺🇸 United States
active

Most cost-efficient Gemini model. Great throughput for high-volume applications.

Context window1000K tokens
Input / 1M tokens$0.075
Output / 1M tokens$0.3

Version History

gemini-2.0-flash-lite-001major

Gemini 2.0 Flash-Lite reaches GA as the most cost-efficient Gemini model. Replaces 1.5 Flash for high-volume workloads.