Qwen3.8-Flash — API Pricing & Benchmarks

Qwen3.8-Flash is Alibaba Qwen's mid reasoning model, released in 2026-08. The API costs $0.15 per 1M input tokens and $0.47 per 1M output tokens ($0.23 blended at a 3:1 ratio), with cached input at $0.01. It supports a 1M-token context window. On the independently measured Artificial Analysis Intelligence Index it scores 55.8, ranking #19 of 68 models tracked on this site, with a median output speed of 86 tokens/sec.

Explore provider routes and additional benchmark sources for Qwen3.8-Flash

ProviderAlibaba Qwen
Input price / 1M tokens$0.15
Output price / 1M tokens$0.47
Cached input / 1M tokens$0.01
Blended price (3:1)$0.23
Context window1M
Max output tokens
Knowledge cutoff
Licenseproprietary
Input modalitiesT
Output speed (tok/s)86
Released2026-08
Artificial Analysis Intelligence Index55.8
GPQA Diamond92.3%
Humanity's Last Exam38%

Where Qwen3.8-Flash fits

On a blended 3:1 basis Qwen3.8-Flash costs $0.23 per 1M tokens, cheaper than 95% of the 80 priced models in this index. It sits in the mid-tier bracket where most production chat, RAG and summarization workloads run: meaningfully cheaper than flagships while keeping most of their capability. Cached input is priced at $0.01 (90% below standard input), which rewards prompt structures with long stable prefixes — see our caching guide.

Price history

PeriodInput $/1MOutput $/1MCached $/1M
2026-08-26 → current$0.15$0.47$0.01

No list-price changes recorded since tracking began. Full dataset: price-history.json (CC BY 4.0).

Alternatives at a similar price

About Alibaba Qwen

Alibaba's Qwen line spans from the flagship Qwen Max, which benchmarks near the frontier at a mid-tier price, down to cheap Flash variants. Qwen models are a frequent value pick on price-vs-performance charts.

More Alibaba Qwen models

← Full comparison table (60+ models, live filters)