Browse all AI Gateway models
Every model available on Vercel AI Gateway, with API access, pricing, and a playground. 328 models · Page 6 of 6.
Search and filter all models →- OpenAITTS-1TTS is a model that converts text to natural sounding spoken text.
- OpenAITTS-1 HDTTS is a model that converts text to natural sounding spoken text. The tts-1-hd model is optimized for high quality text-to-speech use cases.
- GoogleVeo 3.0Veo 3 is designed to handle a range of video generation tasks, from cinematic narratives to dynamic character animations. With Veo 3, you can create more immersive experiences by not only generating stunning visuals, but also audio like dialogue and sound effects.
- GoogleVeo 3.0 Fast GenerateVeo 3 Fast is a quicker and more cost effective version of Veo 3, allowing developers to create videos with sound while maintaining high quality and optimizing for speed and business use cases. Veo 3 Fast offers both text-to-video and image-to-video modalities.
- GoogleVeo 3.1Veo 3.1 is Google's state-of-the-art model for generating high-fidelity, 8-second 720p, 1080p or 4k videos featuring stunning realism and natively generated audio.
- GoogleVeo 3.1 Fast GenerateVeo 3.1 Fast is a specialized, high-speed variant of Google DeepMind’s Veo 3.1 text-to-video model, optimized for rapid generation of 8-second, high-fidelity videos. It is designed to create cinematic, 1080p, or 720p content with improved prompt adherence and native audio, making it ideal for creating quick, high-quality video clips, social media content, and ad creatives.
- GoogleVeo 3.1 Lite GenerateVeo 3.1 Lite Preview is a high-efficiency, developer-first video model providing high-fidelity video generation, editing, and cinematic control. It leverages the state-of-the-art Veo 3.1 model to democratize professional-grade video AI by offering a scalable, programmable interface for creators and enterprises.
- Voyage AIVoyage 3.5Voyage AI's embedding model optimized for general-purpose and multilingual retrieval quality.
- Voyage AIVoyage 3.5 LiteVoyage AI's embedding model optimized for latency and cost.
- Voyage AIVoyage 4Optimized for general-purpose and multilingual retrieval quality. All embeddings created with the 4 series are compatible with each other.
- Voyage AIVoyage 4 LargeThe best general-purpose and multilingual retrieval quality. All embeddings created with the 4 series are compatible with each other.
- Voyage AIVoyage 4 LiteOptimized for latency and cost. All embeddings created with the 4 series are compatible with each other.
- Voyage AIVoyage Code 2Voyage AI's embedding model optimized for code retrieval (17% better than alternatives). This is the previous generation of code embeddings models.
- Voyage AIVoyage Code 3Voyage AI's embedding model optimized for code retrieval.
- Voyage AIVoyage Finance 2Voyage AI's embedding model optimized for finance retrieval and RAG.
- Voyage AIVoyage Law 2Voyage AI's embedding model optimized for legal retrieval and RAG.
- Voyage AIVoyage Rerank 2.5A generalist reranker optimized for quality with instruction-following and multilingual support.
- Voyage AIVoyage Rerank 2.5 LiteA generalist reranker optimized for both latency and quality with instruction-following and multilingual support.
- Voyage AIvoyage-3-largeVoyage AI's embedding model with the best general-purpose and multilingual retrieval quality.
- Alibaba CloudWan v2.5 Text-to-Video Preview
- Alibaba CloudWan v2.6 Image-to-Video
- Alibaba CloudWan v2.6 Image-to-Video Flash
- Alibaba CloudWan v2.6 Reference-to-Video
- Alibaba CloudWan v2.6 Reference-to-Video Flash
- Alibaba CloudWan v2.6 Text-to-Video
- Alibaba CloudWan v2.7 Reference-to-Video
- Alibaba CloudWan v2.7 Text-to-Video
- OpenAIWhisperWhisper is a general-purpose speech recognition model, trained on a large dataset of diverse audio. You can also use it as a multitask model to perform multilingual speech recognition as well as speech translation and language identification.