Gemini 3.6 Flash — API Pricing & Benchmarks
Gemini 3.6 Flash is Google's previous-generation reasoning model, released in 2026-07. The API costs $0.75 per 1M input tokens and $3.75 per 1M output tokens ($1.50 blended at a 3:1 ratio), with cached input at $0.07. It supports a 1.0M-token context window. On the independently measured Artificial Analysis Intelligence Index it scores 51.6, ranking #29 of 68 models tracked on this site, with a median output speed of 196 tokens/sec. Pricing was cut on 2026-08-25 (previously $1.50/$7.50 per 1M tokens).
Explore provider routes and additional benchmark sources for Gemini 3.6 Flash
| Provider | |
| Input price / 1M tokens | $0.75 |
| Output price / 1M tokens | $3.75 |
| Cached input / 1M tokens | $0.07 |
| Blended price (3:1) | $1.50 |
| Context window | 1.0M |
| Max output tokens | 66K |
| Knowledge cutoff | 2026-03 |
| License | proprietary |
| Input modalities | T+I+A+V |
| Output speed (tok/s) | 196 |
| Released | 2026-07 |
| Artificial Analysis Intelligence Index | 51.6 |
| LMArena Elo | 1476 |
| GPQA Diamond | 92.8% |
| Humanity's Last Exam | 40.8% |
Where Gemini 3.6 Flash fits
On a blended 3:1 basis Gemini 3.6 Flash costs $1.50 per 1M tokens, cheaper than 56% of the 80 priced models in this index. It is a previous-generation model: usually only worth choosing if you have an existing integration, negotiated pricing, or need its specific behavior, since newer siblings offer better price-performance. Measured output speed of 196 tokens/sec is well above the index median of 104, making it suitable for interactive, user-facing use. Cached input is priced at $0.07 (90% below standard input), which rewards prompt structures with long stable prefixes — see our caching guide.
Price history
| Period | Input $/1M | Output $/1M | Cached $/1M |
|---|---|---|---|
| 2026-07-26 → 2026-08-25 | $1.50 | $7.50 | $0.15 |
| 2026-08-25 → current | $0.75 | $3.75 | $0.07 |
List-price changes recorded by this site. The full dataset is public: price-history.json (CC BY 4.0).
Alternatives at a similar price
- Gemini 3.8 Flash (Google) — $1.50/1M blended, AA Index 58.7
- Gemini 3.7 Flash (Google) — $1.50/1M blended, AA Index 56
- HyperCLOVA X HCX-007 (Naver) — $1.47/1M blended
- GLM-5 (Zhipu AI (Z.ai)) — $1.55/1M blended, AA Index 40.6
About Google
Google's Gemini models are natively multimodal — most accept text, images, audio and video input — with 1M-token context windows across the range. The Flash line now ships on a roughly monthly cadence (3.6, 3.7, 3.8 Flash between July and September 2026) at the same introductory price, competing on price-performance and speed rather than peak capability.