Gemini 2.0 Flash-Lite — API Pricing & Benchmarks

Gemini 2.0 Flash-Lite is Google's retired model, released in 2025-02 and withdrawn from the API in 2026-06. Its final list price was $0.07 per 1M input tokens and $0.30 per 1M output tokens, with a 1.0M-token context window. This page is kept as a historical reference.

Explore provider routes and additional benchmark sources for Gemini 2.0 Flash-Lite

ProviderGoogle
Statusretired 2026-06 — no longer sold
Input price / 1M tokens$0.07
Output price / 1M tokens$0.30
Cached input / 1M tokens
Blended price (3:1)$0.13
Context window1.0M
Max output tokens8K
Knowledge cutoff2024-08
Licenseproprietary
Input modalitiesT+I+A+V
Output speed (tok/s)
Released2025-02
Artificial Analysis Intelligence Index8.6
GPQA Diamond53.5%
Humanity's Last Exam3.4%

Where Gemini 2.0 Flash-Lite fits

Google retired Gemini 2.0 Flash-Lite in 2026-06; it can no longer be purchased through the standard API, and the prices on this page are its final list prices, kept for reference. If you're migrating an integration, the alternatives below are current models at a similar price point.

Price history

PeriodInput $/1MOutput $/1MCached $/1M
2026-08-25 → current$0.07$0.30

No list-price changes recorded since tracking began. Full dataset: price-history.json (CC BY 4.0).

Alternatives at a similar price

About Google

Google's Gemini models are natively multimodal — most accept text, images, audio and video input — with 1M-token context windows across the range. The Flash line now ships on a roughly monthly cadence (3.6, 3.7, 3.8 Flash between July and September 2026) at the same introductory price, competing on price-performance and speed rather than peak capability.

More Google models

← Full comparison table (60+ models, live filters)