GPT-5.4 — API Pricing & Benchmarks

GPT-5.4 is OpenAI's previous-generation reasoning model, released in 2026-03. The API costs $2.50 per 1M input tokens and $15.0 per 1M output tokens ($5.63 blended at a 3:1 ratio), with cached input at $0.25. It supports a 1.1M-token context window. On the independently measured Artificial Analysis Intelligence Index it scores 53.1, ranking #24 of 68 models tracked on this site, with a median output speed of 150 tokens/sec.

Explore provider routes and additional benchmark sources for GPT-5.4

ProviderOpenAI
Input price / 1M tokens$2.50
Output price / 1M tokens$15.0
Cached input / 1M tokens$0.25
Blended price (3:1)$5.63
Context window1.1M
Max output tokens128K
Knowledge cutoff2025-08
Licenseproprietary
Input modalitiesT+I
Output speed (tok/s)150
Released2026-03
Artificial Analysis Intelligence Index53.1
LMArena Elo1470
GPQA Diamond92%
SWE-bench Verified76.9%
Humanity's Last Exam43.7%

Where GPT-5.4 fits

On a blended 3:1 basis GPT-5.4 costs $5.63 per 1M tokens, cheaper than 19% of the 80 priced models in this index. It is a previous-generation model: usually only worth choosing if you have an existing integration, negotiated pricing, or need its specific behavior, since newer siblings offer better price-performance. Cached input is priced at $0.25 (90% below standard input), which rewards prompt structures with long stable prefixes — see our caching guide.

Price history

PeriodInput $/1MOutput $/1MCached $/1M
2026-07-18 → current$2.50$15.0$0.25

No list-price changes recorded since tracking began. Full dataset: price-history.json (CC BY 4.0).

Alternatives at a similar price

About OpenAI

OpenAI builds the GPT series. GPT-6 Astra (September 2026) sits at the top at $10/$50, and beneath it the GPT-5.6 generation ships in three sizes — Sol, Terra and Luna — that share the same 1.05M-token context window and tooling, so teams can move workloads between price points without changing integration code.

More OpenAI models

← Full comparison table (60+ models, live filters)