GPT-4.1 mini — API Pricing & Benchmarks
GPT-4.1 mini is OpenAI's previous-generation model, released in 2025-04. The API costs $0.40 per 1M input tokens and $1.60 per 1M output tokens ($0.70 blended at a 3:1 ratio), with cached input at $0.10. It supports a 1.0M-token context window. On the independently measured Artificial Analysis Intelligence Index it scores 14.8, ranking #61 of 68 models tracked on this site.
Explore provider routes and additional benchmark sources for GPT-4.1 mini
| Provider | OpenAI |
| Input price / 1M tokens | $0.40 |
| Output price / 1M tokens | $1.60 |
| Cached input / 1M tokens | $0.10 |
| Blended price (3:1) | $0.70 |
| Context window | 1.0M |
| Max output tokens | 33K |
| Knowledge cutoff | 2024-06 |
| License | proprietary |
| Input modalities | T+I |
| Output speed (tok/s) | — |
| Released | 2025-04 |
| Artificial Analysis Intelligence Index | 14.8 |
| LMArena Elo | 1340 |
| GPQA Diamond | 66.4% |
| Humanity's Last Exam | 5% |
Where GPT-4.1 mini fits
On a blended 3:1 basis GPT-4.1 mini costs $0.70 per 1M tokens, cheaper than 73% of the 80 priced models in this index. It is a previous-generation model: usually only worth choosing if you have an existing integration, negotiated pricing, or need its specific behavior, since newer siblings offer better price-performance. Cached input is priced at $0.10 (75% below standard input), which rewards prompt structures with long stable prefixes — see our caching guide.
Price history
| Period | Input $/1M | Output $/1M | Cached $/1M |
|---|---|---|---|
| 2026-08-25 → current | $0.40 | $1.60 | $0.10 |
No list-price changes recorded since tracking began. Full dataset: price-history.json (CC BY 4.0).
Alternatives at a similar price
- Qwen3.7-Plus (Alibaba Qwen) — $0.70/1M blended, AA Index 39.4
- DeepSeek V4 Flash (DeepSeek) — $0.66/1M blended, AA Index 51.8
- Mistral Large 3 (Mistral AI) — $0.75/1M blended, AA Index 15.9
- Gemini 3.5 Flash-Lite (Google) — $0.85/1M blended, AA Index 37.4
About OpenAI
OpenAI builds the GPT series. GPT-6 Astra (September 2026) sits at the top at $10/$50, and beneath it the GPT-5.6 generation ships in three sizes — Sol, Terra and Luna — that share the same 1.05M-token context window and tooling, so teams can move workloads between price points without changing integration code.