This model is 19 months old. A newer version of it has since shipped: Gemini 3.8 Flash (2026-09-02). The details below are still accurate for this model — it just is not what we would recommend today.
Gemini 2.0 Flash-Lite
Google DeepMind🇺🇸 United States
Most cost-efficient Gemini model. Great throughput for high-volume applications.
Context window1000K tokens
Input / 1M tokens$0.075
Output / 1M tokens$0.3
Version History
gemini-2.0-flash-lite-001major
Gemini 2.0 Flash-Lite reaches GA as the most cost-efficient Gemini model. Replaces 1.5 Flash for high-volume workloads.