# Browse all AI Gateway models

Every model available on Vercel AI Gateway, with API access, pricing, and a playground. 341 models · Page 6 of 6.

[Search and filter all models →](/ai-gateway/models)

- [Tencent CloudTencent Hy-MT2-PlusHy-MT2-Plus is Tencent Cloud Hunyuan’s 7B translation model, offering an 8K-token context window and broad multilingual coverage. Optimized for translation across 33 languages and five ethnic Chinese and dialect variants, it delivers industry-leading performance on benchmarks including FLORES-200 and WMT25. It excels in professional domains and real-world business scenarios, with strong instruction-following for structured, delimiter-preserving, context-aware, glossary-guided, and style-specific translation.](/ai-gateway/models/hy-mt2-plus)
- [Tencent CloudTencent Hy-MT2-ProHy-MT2-Pro is Tencent Cloud Hunyuan’s flagship translation model, featuring 30B-A3B parameters and an 8K-token context window. It delivers bidirectional translation across 33 core language pairs, plus five minority-language and dialect pairs. With industry-leading performance on FLORES-200 and WMT25, it excels in specialized domains and real-world business scenarios while supporting structured, contextual, terminology-aware, style-specific, and delimiter-preserving translation.](/ai-gateway/models/hy-mt2-pro)
- [Tencent CloudTencent Hy4 PreviewTencent Hy4 preview is Tencent Hy’s open-source large language model, featuring 770B total parameters, 49B active parameters, and a context window exceeding 1 million tokens. Built for real-world productivity, it supports long-horizon coding, cross-document analysis, office content creation, game development, and scientific reasoning, making it well-suited to complex agents and multi-step professional workflows.](/ai-gateway/models/hy4-preview)
- [GoogleText Embedding 005English-focused text embedding model optimized for code and English language tasks.](/ai-gateway/models/text-embedding-005)
- [GoogleText Multilingual Embedding 002Multilingual text embedding model optimized for cross-lingual tasks across many languages.](/ai-gateway/models/text-multilingual-embedding-002)
- [OpenAItext-embedding-3-largeOpenAI's most capable embedding model for both english and non-english tasks.](/ai-gateway/models/text-embedding-3-large)
- [OpenAItext-embedding-3-smallOpenAI's improved, more performant version of their ada embedding model.](/ai-gateway/models/text-embedding-3-small)
- [OpenAItext-embedding-ada-002OpenAI's legacy text embedding model.](/ai-gateway/models/text-embedding-ada-002)
- [AmazonTitan Text Embeddings V2Amazon Titan Text Embeddings V2 is a light weight, efficient multilingual embedding model supporting 1024, 512, and 256 dimensions.](/ai-gateway/models/titan-embed-text-v2)
- [Fish AudioTranscribe-1Turn spoken audio into accurate text — with timed segments — using Fish Audio’s ASR model. Send an audio file, get back the transcript, its duration, and timestamped segments.](/ai-gateway/models/transcribe-1)
- [Arcee AITrinity Large ThinkingTrinity-Large-Thinking is a reasoning-optimized variant of Arcee AI's Trinity-Large family — a 398B-parameter sparse Mixture-of-Experts (MoE) model with approximately 13B active parameters per token. Built on Trinity-Large-Base and post-trained with extended chain-of-thought reasoning and agentic RL, Trinity-Large-Thinking delivers state-of-the-art performance on agentic benchmarks while maintaining strong general capabilities.](/ai-gateway/models/trinity-large-thinking)
- [OpenAITTS-1TTS is a model that converts text to natural sounding spoken text.](/ai-gateway/models/tts-1)
- [OpenAITTS-1 HDTTS is a model that converts text to natural sounding spoken text. The tts-1-hd model is optimized for high quality text-to-speech use cases.](/ai-gateway/models/tts-1-hd)
- [GoogleVeo 3.0Veo 3 is designed to handle a range of video generation tasks, from cinematic narratives to dynamic character animations. With Veo 3, you can create more immersive experiences by not only generating stunning visuals, but also audio like dialogue and sound effects.](/ai-gateway/models/veo-3.0-generate-001)
- [GoogleVeo 3.0 Fast GenerateVeo 3 Fast is a quicker and more cost effective version of Veo 3, allowing developers to create videos with sound while maintaining high quality and optimizing for speed and business use cases. Veo 3 Fast offers both text-to-video and image-to-video modalities.](/ai-gateway/models/veo-3.0-fast-generate-001)
- [GoogleVeo 3.1Veo 3.1 is Google's state-of-the-art model for generating high-fidelity, 8-second 720p, 1080p or 4k videos featuring stunning realism and natively generated audio.](/ai-gateway/models/veo-3.1-generate-001)
- [GoogleVeo 3.1 Fast GenerateVeo 3.1 Fast is a specialized, high-speed variant of Google DeepMind’s Veo 3.1 text-to-video model, optimized for rapid generation of 8-second, high-fidelity videos. It is designed to create cinematic, 1080p, or 720p content with improved prompt adherence and native audio, making it ideal for creating quick, high-quality video clips, social media content, and ad creatives.](/ai-gateway/models/veo-3.1-fast-generate-001)
- [GoogleVeo 3.1 Lite GenerateVeo 3.1 Lite Preview is a high-efficiency, developer-first video model providing high-fidelity video generation, editing, and cinematic control. It leverages the state-of-the-art Veo 3.1 model to democratize professional-grade video AI by offering a scalable, programmable interface for creators and enterprises.](/ai-gateway/models/veo-3.1-lite-generate-001)
- [Voyage AI by MongoDBVoyage 3.5Voyage AI's embedding model optimized for general-purpose and multilingual retrieval quality.](/ai-gateway/models/voyage-3.5)
- [Voyage AI by MongoDBVoyage 3.5 LiteVoyage AI's embedding model optimized for latency and cost.](/ai-gateway/models/voyage-3.5-lite)
- [Voyage AI by MongoDBVoyage 4Optimized for general-purpose and multilingual retrieval quality. All embeddings created with the 4 series are compatible with each other.](/ai-gateway/models/voyage-4)
- [Voyage AI by MongoDBVoyage 4 LargeThe best general-purpose and multilingual retrieval quality. All embeddings created with the 4 series are compatible with each other.](/ai-gateway/models/voyage-4-large)
- [Voyage AI by MongoDBVoyage 4 LiteOptimized for latency and cost. All embeddings created with the 4 series are compatible with each other.](/ai-gateway/models/voyage-4-lite)
- [Voyage AI by MongoDBVoyage Code 2Voyage AI's embedding model optimized for code retrieval (17% better than alternatives). This is the previous generation of code embeddings models.](/ai-gateway/models/voyage-code-2)
- [Voyage AI by MongoDBVoyage Code 3Voyage AI's embedding model optimized for code retrieval.](/ai-gateway/models/voyage-code-3)
- [Voyage AI by MongoDBVoyage Finance 2Voyage AI's embedding model optimized for finance retrieval and RAG.](/ai-gateway/models/voyage-finance-2)
- [Voyage AI by MongoDBVoyage Law 2Voyage AI's embedding model optimized for legal retrieval and RAG.](/ai-gateway/models/voyage-law-2)
- [Voyage AI by MongoDBVoyage Rerank 2.5A generalist reranker optimized for quality with instruction-following and multilingual support.](/ai-gateway/models/rerank-2.5)
- [Voyage AI by MongoDBVoyage Rerank 2.5 LiteA generalist reranker optimized for both latency and quality with instruction-following and multilingual support.](/ai-gateway/models/rerank-2.5-lite)
- [Voyage AI by MongoDBvoyage-3-largeVoyage AI's embedding model with the best general-purpose and multilingual retrieval quality.](/ai-gateway/models/voyage-3-large)
- [Alibaba CloudWan v2.5 Text-to-Video Preview](/ai-gateway/models/wan-v2.5-t2v-preview)
- [Alibaba CloudWan v2.6 Image-to-Video](/ai-gateway/models/wan-v2.6-i2v)
- [Alibaba CloudWan v2.6 Image-to-Video Flash](/ai-gateway/models/wan-v2.6-i2v-flash)
- [Alibaba CloudWan v2.6 Reference-to-Video](/ai-gateway/models/wan-v2.6-r2v)
- [Alibaba CloudWan v2.6 Reference-to-Video Flash](/ai-gateway/models/wan-v2.6-r2v-flash)
- [Alibaba CloudWan v2.6 Text-to-Video](/ai-gateway/models/wan-v2.6-t2v)
- [Alibaba CloudWan v2.7 Reference-to-Video](/ai-gateway/models/wan-v2.7-r2v)
- [Alibaba CloudWan v2.7 Text-to-Video](/ai-gateway/models/wan-v2.7-t2v)
- [Alibaba CloudWan v3.0 VideoAll-in-one video generation model supporting text-to-video, image-to-video, first/last-frame, and omni-modal reference-based generation (image, video, and audio references) with synchronized audio, at up to 30 seconds per clip.](/ai-gateway/models/wan-v3.0-video)
- [Alibaba CloudWan v3.0 Video PrimeWan 3.0 Video Prime is Alibaba’s high-speed, all-in-one AI video generation model for creating polished clips from text, images, video, and audio references. It produces videos up to 30 seconds long with synchronized dialogue, music, and sound effects, while accelerated generation makes it ideal for rapid creative iteration and production workflows.](/ai-gateway/models/wan-v3.0-video-prime)
- [OpenAIWhisperWhisper is a general-purpose speech recognition model, trained on a large dataset of diverse audio. You can also use it as a multitask model to perform multilingual speech recognition as well as speech translation and language identification.](/ai-gateway/models/whisper-1)