[OpenAI](/ai-gateway/models/labs/openai)

# GPT-4o mini Transcribe

GPT-4o mini Transcribe is a speech-to-text model built on the GPT-4o mini architecture, delivering lower word error rates and better language recognition than the original Whisper models at the cost-efficient end of OpenAI's transcription lineup. Your use is subject to OpenAI's [Terms](https://openai.com/policies/terms-of-use) & [Privacy](https://openai.com/policies/privacy-policy) Policies.

[Use with AI Gateway](https://vercel.com/d?to=%2F%5Bteam%5D%2F%7E%2Fai%3Futm_source%3Dgateway-model-page%26utm_campaign%3Dai-gateway-models&title=Get+Started+with+Vercel+AI+Gateway) [View docs](https://vercel.com/docs/ai-gateway)

```
1import { experimental_transcribe as transcribe } from 'ai';
2import { gateway } from '@ai-sdk/gateway';
3import { readFile } from 'node:fs/promises';
4

5const result = await transcribe({
6  model: gateway.transcriptionModel('openai/gpt-4o-mini-transcribe'),
7  audio: await readFile('audio.mp3'),
8});
```

[Read docs](https://vercel.com/docs/ai-gateway/sdks-and-apis/ai-sdk)

[Overview](/ai-gateway/models/gpt-4o-mini-transcribe) [About](/ai-gateway/models/gpt-4o-mini-transcribe/about) [Providers](/ai-gateway/models/gpt-4o-mini-transcribe/providers) [Similar](/ai-gateway/models/gpt-4o-mini-transcribe/similar) [FAQ](/ai-gateway/models/gpt-4o-mini-transcribe/faq)

## [Copy link to heading](#playground)Playground

Try out GPT-4o mini Transcribe by OpenAI. Usage is billed to your team at API rates. Free users (those who haven't made a payment) get $5 of credits every 30 days.

![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75)GPT-4o mini Transcribe

### Speech to text

Record a short clip from your microphone and the model transcribes it to text.

Idle

![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=96&q=75)

Record a clip to see the transcript here.

## [Copy link to heading](#providers)Providers

Route requests across multiple providers. Copy a provider slug to set your preference. Visit the [docs](/docs/ai-gateway/provider-options) for more info. Using a provider means you agree to their terms, listed under Legal.

| Provider |
| --- |

| Input | Capabilities | ZDR | No Training | Release Date |  |
| --- | --- | --- | --- | --- | --- |

| ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) [OpenAI](/ai-gateway/models/providers/openai) Legal:[Terms](https://openai.com/policies/terms-of-use)•[Privacy](https://openai.com/policies/privacy-policy) |  |  |  |  | 03/13/2024 |  |
| --- | --- | --- | --- | --- | --- | --- |

## [Copy link to heading](#more-models-by-openai)More models by OpenAI

All

Text

Code

| Model |
| --- |

| Context | Latency | Throughput | Input | Output | Cache | Web Search | Capabilities | Providers | ZDR | No Training | Release Date |  |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- |

| ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) [openai/gpt-5.6-luna](/ai-gateway/models/gpt-5.6-luna) | 1.1M | 0.9s | 141tps | $1/M$0.20/M Fast $0.40/M+1 more | $6/M$1.20/M Fast $2.40/M+1 more | Read: $0.1/M$0.02/M+1 more Write: $1.25/M$0.25/M+1 more | $10/K \+ input costs | +4 | ![azure logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fazure.png&w=48&q=75) ![bedrock logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Famazon%2520bedrock.png&w=48&q=75) ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) |  |  | 07/09/2026 |  |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- |
| ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) [openai/gpt\-5.6-sol](/ai-gateway/models/gpt-5.6-sol) | 1.1M | 4.2s | 111tps | $4/M$2/M Fast $4/M+1 more | $20/M$10/M Fast $20/M+1 more | Read: $0.4/M$0.2/M+1 more Write: $5/M$2.5/M+1 more | $10/K \+ input costs | +4 | ![azure logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fazure.png&w=48&q=75) ![bedrock logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Famazon%2520bedrock.png&w=48&q=75) ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) |  |  | 07/09/2026 |  |
| ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) [openai/gpt-5.6-terra](/ai-gateway/models/gpt-5.6-terra) | 1.1M | 2.0s | 191tps | $2.50/M$2/M Fast $4/M+1 more | $15/M$12/M Fast $24/M+1 more | Read: $0.25/M$0.2/M+1 more Write: $3.13/M$2.5/M+1 more | $10/K \+ input costs | +4 | ![azure logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fazure.png&w=48&q=75) ![bedrock logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Famazon%2520bedrock.png&w=48&q=75) ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) |  |  | 07/09/2026 |  |
| ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) [openai/gpt-5.4](/ai-gateway/models/gpt-5.4) | 1.1M | 0.7s | 80tps | $2.50/MFast $5/M+1 more | $15/MFast $30/M+1 more | Read: $0.25/M+1 more Write: — | $10/K \+ input costs | +4 | ![azure logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fazure.png&w=48&q=75) ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) |  |  | 03/05/2026 |  |
| ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) [openai/gpt-5-nano](/ai-gateway/models/gpt-5-nano) | 400K | 3.2s | 203tps | $0.05/M | $0.40/M | Read:$0.01/M Write:— | $14/K \+ input costs | +3 | ![azure logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fazure.png&w=48&q=75) ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) |  |  | 08/07/2025 |  |
| ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) [openai/gpt-5-mini](/ai-gateway/models/gpt-5-mini) | 400K | 2.9s | 94tps | $0.25/MFast $0.45/M | $2/MFast $3.60/M | Read:$0.03/M Write:— | $14/K \+ input costs | +3 | ![azure logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fazure.png&w=48&q=75) ![openai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fopenai.png&w=48&q=75) |  |  | 08/07/2025 |  |

## [Copy link to heading](#about-gpt-4o-mini-transcribe)About GPT-4o mini Transcribe

GPT-4o mini Transcribe launched on March 13, 2024 as part of OpenAI's next-generation audio models for the API. The model builds on the GPT-4o mini architecture and was pretrained on specialized audio-centric datasets. OpenAI used enhanced distillation techniques to transfer knowledge from larger audio models into this smaller one.

Compared to the original Whisper models, GPT-4o mini Transcribe improves word error rate and language recognition. OpenAI attributes the gains to reinforcement learning work and midtraining on diverse, high-quality audio data. The improvements show up most in difficult conditions: accents, noisy environments, and varying speech speeds.

Through AI Gateway, GPT-4o mini Transcribe handles transcription jobs behind the same authentication, observability, and spend controls as your text models. Send audio with the AI SDK's `transcribe` function and get back the transcript, plus segments and language metadata where available.

## [Copy link to heading](#what-to-consider-when-choosing-a-provider)What To Consider When Choosing a Provider

- Configuration: Audio support on AI Gateway is in beta. You call GPT-4o mini Transcribe with the AI SDK's `transcribe` function, passing audio as a file buffer, base64 string, or URL. If accuracy on your hardest recordings is the deciding factor, compare results against `gpt-4o-transcribe` on your own audio before settling on the mini variant.
- Zero Data Retention: Zero Data Retention is offered on a per-provider and model basis. See the [documentation](https://vercel.com/docs/ai-gateway/security-and-compliance/zdr) for details.
- Authentication: AI Gateway authenticates requests using an [API key](https://vercel.com/docs/ai-gateway/authentication-and-byok#api-key-authentication) or [OIDC token](https://vercel.com/docs/ai-gateway/authentication-and-byok#oidc-token-authentication). You do not need to manage provider credentials directly.

## [Copy link to heading](#when-to-use-gpt-4o-mini-transcribe)When to Use GPT-4o mini Transcribe

### Best for

- High-volume transcription pipelines: Per-request cost shapes the architecture and the mini rate keeps unit economics workable
- Call and meeting transcription: Accuracy improvements over Whisper-generation models reduce cleanup work downstream
- Difficult audio conditions: Accents, background noise, and varying speech speeds that trip up older models
- Unified gateway workflows: Adding speech input to apps that already route text models through AI Gateway

### Consider alternatives when

- Maximum transcription accuracy: `gpt-4o-transcribe` is the stronger variant when difficult audio justifies a higher rate
- Speech translation needs: `whisper-1` handles translation to English and language identification as a multitask model
- Live voice conversations: The gpt-realtime family serves speech-to-speech agents rather than transcription jobs

## [Copy link to heading](#conclusion)Conclusion

GPT-4o mini Transcribe is a practical default for transcription through AI Gateway: more accurate than Whisper-generation models and priced for volume. Start here for most speech-to-text workloads, and step up to `gpt-4o-transcribe` when your hardest audio demands it.