Qwen3.8-Flash — API Pricing & Benchmarks
Qwen3.8-Flash is Alibaba Qwen's mid reasoning model, released in 2026-08. The API costs $0.15 per 1M input tokens and $0.47 per 1M output tokens ($0.23 blended at a 3:1 ratio), with cached input at $0.01. It supports a 1M-token context window. On the independently measured Artificial Analysis Intelligence Index it scores 55.8, ranking #19 of 68 models tracked on this site, with a median output speed of 86 tokens/sec.
Explore provider routes and additional benchmark sources for Qwen3.8-Flash
| Provider | Alibaba Qwen |
| Input price / 1M tokens | $0.15 |
| Output price / 1M tokens | $0.47 |
| Cached input / 1M tokens | $0.01 |
| Blended price (3:1) | $0.23 |
| Context window | 1M |
| Max output tokens | — |
| Knowledge cutoff | — |
| License | proprietary |
| Input modalities | T |
| Output speed (tok/s) | 86 |
| Released | 2026-08 |
| Artificial Analysis Intelligence Index | 55.8 |
| GPQA Diamond | 92.3% |
| Humanity's Last Exam | 38% |
Where Qwen3.8-Flash fits
On a blended 3:1 basis Qwen3.8-Flash costs $0.23 per 1M tokens, cheaper than 95% of the 80 priced models in this index. It sits in the mid-tier bracket where most production chat, RAG and summarization workloads run: meaningfully cheaper than flagships while keeping most of their capability. Cached input is priced at $0.01 (90% below standard input), which rewards prompt structures with long stable prefixes — see our caching guide.
Price history
| Period | Input $/1M | Output $/1M | Cached $/1M |
|---|---|---|---|
| 2026-08-26 → current | $0.15 | $0.47 | $0.01 |
No list-price changes recorded since tracking began. Full dataset: price-history.json (CC BY 4.0).
Alternatives at a similar price
- GLM-5.3-Flash (Zhipu AI (Z.ai)) — $0.24/1M blended, AA Index 57.5
- GPT-4o mini (OpenAI) — $0.26/1M blended, AA Index 6.7
- Mistral Small 4 (Mistral AI) — $0.26/1M blended, AA Index 19.7
- Solar Pro 3 (Upstage) — $0.26/1M blended, AA Index 14.5
About Alibaba Qwen
Alibaba's Qwen line spans from the flagship Qwen Max, which benchmarks near the frontier at a mid-tier price, down to cheap Flash variants. Qwen models are a frequent value pick on price-vs-performance charts.