Skip to content
Dashboard

GPT 5.4

GPT 5.4 is the standard tier of the GPT-5.4 model family, extending the agentic and reasoning capabilities of GPT-5.3 Codex to all domains including knowledge work, multi-step workflows, and analysis. Your use is subject to OpenAI's Terms & Privacy Policies.

ReasoningTool UseVision (Image)File InputImplicit CachingWeb SearchHas Fast ModeWebsockets
import { streamText } from 'ai'
const result = streamText({
model: 'openai/gpt-5.4',
prompt: 'Why is the sky blue?'
})
Read docs

Copy link to headingPlayground

Try out GPT 5.4 by OpenAI. Usage is billed to your team at API rates. Free users (those who haven't made a payment) get $5 of credits every 30 days.

openai logo
openai logo

GPT 5.4

Copy link to headingProviders

Route requests across multiple providers. Copy a provider slug to set your preference. Visit the docs for more info. Using a provider means you agree to their terms, listed under Legal.

Provider
Context
Max Output
Latency
Throughput
Input
Output
Cache
Web Search
Capabilities
ZDR
No Training
Regional Inference
Release Date
1.1M128K
0.8s
76tps
$2.50/M+4 more
$15/M+4 more
Read:$0.25/M+2 more
Write:
$10/K
+ input costs
+4
US
03/05/2026
1M128K
3.6s
52tps
$2.50/M+1 more
$15/M+1 more
Read:
$0.25/M+1 more
Write:
$14/K
+ input costs
+3
03/05/2026

Copy link to headingThroughput

P50 throughput on live AI Gateway traffic, in tokens per second (TPS). Visit the docs for more info.

Copy link to headingLatency

P50 time to first token (TTFT) on live AI Gateway traffic, in milliseconds. View the docs for more info.

Copy link to headingUptime

Direct request success rate on AI Gateway and per-provider. Visit the docs for more info.

Copy link to headingMore models by OpenAI

Model
Context
Latency
Throughput
Input
Output
Cache
Web Search
Capabilities
Providers
ZDR
No Training
Release Date
1.1M
2.2s
148tps
$1/M$0.20/M
+1 more
$6/M$1.20/M
+1 more
Read:
$0.1/M$0.02/M+1 more
Write:
$1.25/M$0.25/M+1 more
$10/K
+ input costs
+4
azure logo
bedrock logo
openai logo
07/09/2026
1.1M
2.3s
121tps
$5/M+1 more
$30/M+1 more
Read:
$0.5/M+1 more
Write:
$6.25/M+1 more
$10/K
+ input costs
+4
azure logo
bedrock logo
openai logo
07/09/2026
1.1M
2.3s
104tps
$2.50/M$2/M
+1 more
$15/M$12/M
+1 more
Read:
$0.25/M$0.2/M+1 more
Write:
$3.13/M$2.5/M+1 more
$10/K
+ input costs
+4
azure logo
bedrock logo
openai logo
07/09/2026
400K
4.7s
192tps
$0.05/M
$0.40/M
Read:$0.01/M
Write:
$14/K
+ input costs
+3
azure logo
openai logo
08/07/2025
400K
3.5s
126tps
$0.25/M
$2/M
Read:$0.03/M
Write:
$14/K
+ input costs
+3
azure logo
openai logo
08/07/2025
131K
0.1s
315tps
$0.35/M
$0.75/M
Read:$0.25/M
Write:
baseten logo
bedrock logo
cerebras logo
+5
08/05/2025

GPT 5.4 became available on March 5, 2026 on AI Gateway as the standard tier of the GPT-5.4 model family. It extends the agentic and reasoning leaps introduced in GPT-5.3 Codex to all domains, including knowledge work like reports, spreadsheets, presentations, and analysis.

The model handles complex multi-step workflows more reliably than previous generations, including tasks that involve tools, research, and pulling from multiple sources. It's faster and more token-efficient than GPT-5.2, delivering better results at lower cost per task.

With the context window of 1.1M tokens and the full API feature set, GPT 5.4 supports text, image, and mixed-modality inputs. If you're starting a new project or upgrading from an earlier GPT-5.x model, it's the default starting point for general-purpose work.

Copy link to headingWhat To Consider When Choosing a Provider

  • Configuration: GPT 5.4 brings the agentic and reasoning leaps from GPT-5.3 Codex into all domains, not just coding.
  • Configuration: It's faster and more token-efficient than GPT-5.2, meaning lower costs and shorter latencies on comparable tasks.
  • Zero Data Retention: AI Gateway supports Zero Data Retention for this model via direct gateway requests (BYOK is not included). To configure this, check the documentation.
  • Authentication: AI Gateway authenticates requests using an API key or OIDC token. You do not need to manage provider credentials directly.

Copy link to headingWhen to Use GPT 5.4

Best for

  • Complex multi-step workflows: Tasks involving tools, research, and pulling from multiple sources
  • Knowledge work: Reports, spreadsheets, presentations, and analysis across business domains
  • Advanced code generation: Strong coding performance with GPT-5.4 generation reasoning improvements
  • Agentic applications: Autonomous agents that coordinate tools, research, and multi-source workflows
  • General-purpose AI: New projects that benefit from GPT-5.4 generation capability

Consider alternatives when

  • Cost optimization: GPT-5.4 mini for production workloads where cost efficiency matters
  • High-volume lightweight tasks: GPT-5.4 nano for classification, routing, and sub-agent workflows
  • Extended reasoning: GPT-5.4 pro for maximum performance on the most complex tasks
  • Pure chain-of-thought: O3 for mathematical and scientific reasoning tasks

GPT 5.4 brings agentic reasoning to all domains with improved speed and efficiency over GPT-5.2, available through AI Gateway. It is the standard tier of the GPT-5.4 family.