[OpenAI](/ai-gateway/models/labs/openai)

# o3

o3 is OpenAI's advanced reasoning model that succeeds o1, delivering stronger chain-of-thought performance on mathematical, scientific, and coding problems with improved efficiency and full tool support. Your use is subject to OpenAI's [Terms](https://openai.com/policies/terms-of-use) & [Privacy](https://openai.com/policies/privacy-policy) Policies.

File InputImplicit CachingReasoningTool UseVision (Image)Web SearchHas Fast Mode

[Use with AI Gateway](https://vercel.com/d?to=%2F%5Bteam%5D%2F%7E%2Fai%3Futm_source%3Dgateway-model-page%26utm_campaign%3Dai-gateway-models&title=Get+Started+with+Vercel+AI+Gateway) [View docs](https://vercel.com/docs/ai-gateway)

AI SDKChat CompletionsMessagesResponses

```
1import { streamText } from 'ai'
2

3const result = streamText({
4  model: 'openai/o3',
5  prompt: 'Why is the sky blue?'
6})
```

[Read docs](https://vercel.com/docs/ai-gateway/sdks-and-apis/ai-sdk)

[Overview](/ai-gateway/models/o3) [API](/ai-gateway/models/o3/api) [About](/ai-gateway/models/o3/about) [Providers](/ai-gateway/models/o3/providers) [Throughput](/ai-gateway/models/o3/throughput) [Latency](/ai-gateway/models/o3/latency) [Uptime](/ai-gateway/models/o3/uptime) [Status](/ai-gateway/models/o3/status) [Similar](/ai-gateway/models/o3/similar) [FAQ](/ai-gateway/models/o3/faq)

## [Copy link to heading](#playground)Playground

Try out o3 by OpenAI. Usage is billed to your team at API rates. Free users (those who haven't made a payment) get $5 of credits every 30 days.

![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75)o3

![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=96&q=75)

o3

## [Copy link to heading](#providers)Providers

Route requests across multiple providers. Copy a provider slug to set your preference. Visit the [docs](/docs/ai-gateway/provider-options) for more info. Using a provider means you agree to their terms, listed under Legal.

| Provider |
| --- |

| Context | Max Output | Latency | Throughput | Input | Output | Cache | Web Search | Capabilities | ZDR | No Training | Release Date |  |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- |

| ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) [OpenAI](/ai-gateway/models/providers/openai) Legal:[Terms](https://openai.com/policies/terms-of-use)•[Privacy](https://openai.com/policies/privacy-policy) | 200K | 100K | 0.6s | 130tps | $2/M+2 more | $8/M+2 more | Read:$0.5/M+2 more Write:— | $10/K \+ input costs | +3 |  |  | 04/16/2025 |  |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- |

## [Copy link to heading](#throughput)Throughput24 hours

1W

1D

P50 throughput on live AI Gateway traffic, in tokens per second (TPS). Visit the [docs](https://vercel.com/docs/ai-gateway/models-and-providers/metrics) for more info.

## [Copy link to heading](#latency)Latency24 hours

1W

1D

P50 time to first token (TTFT) on live AI Gateway traffic, in milliseconds. View the [docs](https://vercel.com/docs/ai-gateway/models-and-providers/metrics) for more info.

## [Copy link to heading](#uptime)Uptime24 hours

1W

1D

1H

Direct request success rate on AI Gateway and per-provider. Visit the [docs](https://vercel.com/docs/ai-gateway/models-and-providers/uptime) for more info.

1W

1D

1H

## [Copy link to heading](#more-models-by-openai)More models by OpenAI

All

Text

Code

| Model |
| --- |

| Context | Latency | Throughput | Input | Output | Cache | Web Search | Capabilities | Providers | ZDR | No Training | Release Date |  |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- |

| ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) [openai/gpt-5.6-luna](/ai-gateway/models/gpt-5.6-luna) | 1.1M | 2.0s | 187tps | $1/M$0.20/M Fast $0.40/M+1 more | $6/M$1.20/M Fast $2.40/M+1 more | Read: $0.1/M$0.02/M+1 more Write: $1.25/M$0.25/M+1 more | $10/K \+ input costs | +4 | ![azure logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fazure.png&w=48&q=75) ![bedrock logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Famazon%2520bedrock.png&w=48&q=75) ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) |  |  | 07/09/2026 |  |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- |
| ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) [openai/gpt\-5.6-sol](/ai-gateway/models/gpt-5.6-sol) | 1.1M | 2.7s | 159tps | $4/M$2/M Fast $4/M+1 more | $20/M$10/M Fast $20/M+1 more | Read: $0.4/M$0.2/M+1 more Write: $5/M$2.5/M+1 more | $10/K \+ input costs | +4 | ![azure logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fazure.png&w=48&q=75) ![bedrock logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Famazon%2520bedrock.png&w=48&q=75) ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) |  |  | 07/09/2026 |  |
| ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) [openai/gpt-5.6-terra](/ai-gateway/models/gpt-5.6-terra) | 1.1M | 1.1s | 187tps | $2.50/M$2/M Fast $4/M+1 more | $15/M$12/M Fast $24/M+1 more | Read: $0.25/M$0.2/M+1 more Write: $3.13/M$2.5/M+1 more | $10/K \+ input costs | +4 | ![azure logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fazure.png&w=48&q=75) ![bedrock logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Famazon%2520bedrock.png&w=48&q=75) ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) |  |  | 07/09/2026 |  |
| ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) [openai/gpt-5.4](/ai-gateway/models/gpt-5.4) | 1.1M | 0.6s | 85tps | $2.50/MFast $5/M+1 more | $15/MFast $30/M+1 more | Read: $0.25/M+1 more Write: — | $10/K \+ input costs | +4 | ![azure logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fazure.png&w=48&q=75) ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) |  |  | 03/05/2026 |  |
| ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) [openai/gpt-5-nano](/ai-gateway/models/gpt-5-nano) | 400K | 4.0s | 230tps | $0.05/M | $0.40/M | Read:$0.01/M Write:— | $14/K \+ input costs | +3 | ![azure logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fazure.png&w=48&q=75) ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) |  |  | 08/07/2025 |  |
| ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) [openai/gpt-5-mini](/ai-gateway/models/gpt-5-mini) | 400K | 3.8s | 147tps | $0.25/MFast $0.45/M | $2/MFast $3.60/M | Read:$0.03/M Write:— | $14/K \+ input costs | +3 | ![azure logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fazure.png&w=48&q=75) ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) |  |  | 08/07/2025 |  |

## [Copy link to heading](#about-o3)About o3

o3 was released on April 16, 2025 as the successor to o1 in OpenAI's reasoning model series. It advances the chain-of-thought paradigm introduced by o1-preview: the model generates internal reasoning tokens, working through problems step by step, checking its work, and trying alternative approaches before producing a final answer.

o3 improves on o1 across key reasoning benchmarks while using reasoning tokens more efficiently. It launches with the full set of production API features: function calling for tool use, structured outputs via JSON schema constrained decoding, vision input for reasoning over images and diagrams, and developer system messages for behavioral control.

The context window of 200K tokens accommodates the lengthy inputs that complex reasoning tasks demand. The model also supports the `reasoning_effort` parameter, letting you control how deeply it thinks on a per-request basis for efficient handling of mixed-difficulty workloads.

## [Copy link to heading](#what-to-consider-when-choosing-a-provider)What To Consider When Choosing a Provider

- Configuration: o3 generates internal reasoning tokens that work through problems step by step before producing a visible response. This trades latency for accuracy on hard problems.
- Configuration: Unlike earlier reasoning previews, o3 ships with function calling, structured outputs, vision, and system messages from day one.
- Zero Data Retention: Zero Data Retention is offered on a per-provider and model basis. See the [documentation](https://vercel.com/docs/ai-gateway/security-and-compliance/zdr) for details.
- Authentication: AI Gateway authenticates requests using an [API key](https://vercel.com/docs/ai-gateway/authentication-and-byok#api-key-authentication) or [OIDC token](https://vercel.com/docs/ai-gateway/authentication-and-byok#oidc-token-authentication). You do not need to manage provider credentials directly.

## [Copy link to heading](#when-to-use-o3)When to Use o3

### Best for

- Advanced mathematical reasoning: Competition-level math, proofs, and quantitative analysis
- Complex coding problems: Algorithm design, optimization, and architectural reasoning
- Scientific analysis: Multi-step derivations in physics, chemistry, and biology
- Agentic reasoning: Agent backbones that need deep deliberation before acting
- Hard problem solving: Any task where extended chain-of-thought produces measurably better results

### Consider alternatives when

- General-purpose tasks: GPT-5 or GPT-5.2 for conversational and generative workloads that don't need chain-of-thought
- Cost-sensitive reasoning: O4-mini for reasoning at a lower price point
- Maximum reasoning compute: O3-pro for the hardest problems that benefit from extended computation
- Fast responses: GPT-5.1 instant or GPT-4o when latency matters more than reasoning depth

## [Copy link to heading](#conclusion)Conclusion

o3 advances reasoning model capability beyond o1, delivering stronger performance and greater efficiency than o1 with full production API features. For the hardest analytical, mathematical, and coding problems routed through AI Gateway, it is the standard reasoning model.