Gemini 3.5 Flash — API Pricing & Benchmarks

Gemini 3.5 Flash is Google's previous-generation reasoning model, released in 2026-05. The API costs $1.50 per 1M input tokens and $9.00 per 1M output tokens ($3.38 blended at a 3:1 ratio), with cached input at $0.15. It supports a 1.0M-token context window. On the independently measured Artificial Analysis Intelligence Index it scores 52, ranking #27 of 68 models tracked on this site, with a median output speed of 251 tokens/sec.

Explore provider routes and additional benchmark sources for Gemini 3.5 Flash

ProviderGoogle
Input price / 1M tokens$1.50
Output price / 1M tokens$9.00
Cached input / 1M tokens$0.15
Blended price (3:1)$3.38
Context window1.0M
Max output tokens66K
Knowledge cutoff2025-01
Licenseproprietary
Input modalitiesT+I+A+V
Output speed (tok/s)251
Released2026-05
Artificial Analysis Intelligence Index52
LMArena Elo1484
GPQA Diamond92.2%
SWE-bench Verified79.3%
Humanity's Last Exam42.7%
MMMU81.2%

Where Gemini 3.5 Flash fits

On a blended 3:1 basis Gemini 3.5 Flash costs $3.38 per 1M tokens, cheaper than 33% of the 80 priced models in this index. It is a previous-generation model: usually only worth choosing if you have an existing integration, negotiated pricing, or need its specific behavior, since newer siblings offer better price-performance. Measured output speed of 251 tokens/sec is well above the index median of 104, making it suitable for interactive, user-facing use. Cached input is priced at $0.15 (90% below standard input), which rewards prompt structures with long stable prefixes — see our caching guide.

Price history

PeriodInput $/1MOutput $/1MCached $/1M
2026-07-18 → current$1.50$9.00$0.15

No list-price changes recorded since tracking began. Full dataset: price-history.json (CC BY 4.0).

Alternatives at a similar price

About Google

Google's Gemini models are natively multimodal — most accept text, images, audio and video input — with 1M-token context windows across the range. The Flash line now ships on a roughly monthly cadence (3.6, 3.7, 3.8 Flash between July and September 2026) at the same introductory price, competing on price-performance and speed rather than peak capability.

More Google models

← Full comparison table (60+ models, live filters)