Gemini 2.0 Flash-Lite — API Pricing & Benchmarks
Gemini 2.0 Flash-Lite is Google's retired model, released in 2025-02 and withdrawn from the API in 2026-06. Its final list price was $0.07 per 1M input tokens and $0.30 per 1M output tokens, with a 1.0M-token context window. This page is kept as a historical reference.
Explore provider routes and additional benchmark sources for Gemini 2.0 Flash-Lite
| Provider | |
| Status | retired 2026-06 — no longer sold |
| Input price / 1M tokens | $0.07 |
| Output price / 1M tokens | $0.30 |
| Cached input / 1M tokens | — |
| Blended price (3:1) | $0.13 |
| Context window | 1.0M |
| Max output tokens | 8K |
| Knowledge cutoff | 2024-08 |
| License | proprietary |
| Input modalities | T+I+A+V |
| Output speed (tok/s) | — |
| Released | 2025-02 |
| Artificial Analysis Intelligence Index | 8.6 |
| GPQA Diamond | 53.5% |
| Humanity's Last Exam | 3.4% |
Where Gemini 2.0 Flash-Lite fits
Google retired Gemini 2.0 Flash-Lite in 2026-06; it can no longer be purchased through the standard API, and the prices on this page are its final list prices, kept for reference. If you're migrating an integration, the alternatives below are current models at a similar price point.
Price history
| Period | Input $/1M | Output $/1M | Cached $/1M |
|---|---|---|---|
| 2026-08-25 → current | $0.07 | $0.30 | — |
No list-price changes recorded since tracking began. Full dataset: price-history.json (CC BY 4.0).
Alternatives at a similar price
- GPT-4.1 nano (OpenAI) — $0.18/1M blended, AA Index 9.6
- Qwen3.8-Flash (Alibaba Qwen) — $0.23/1M blended, AA Index 55.8
- GLM-5.3-Flash (Zhipu AI (Z.ai)) — $0.24/1M blended, AA Index 57.5
- GPT-4o mini (OpenAI) — $0.26/1M blended, AA Index 6.7
About Google
Google's Gemini models are natively multimodal — most accept text, images, audio and video input — with 1M-token context windows across the range. The Flash line now ships on a roughly monthly cadence (3.6, 3.7, 3.8 Flash between July and September 2026) at the same introductory price, competing on price-performance and speed rather than peak capability.