# Browse all AI Gateway models

Every model available on Vercel AI Gateway, with API access, pricing, and a playground. 340 models · Page 5 of 6.

[Search and filter all models →](/ai-gateway/models)

- [Alibaba CloudQwen 3.5 PlusThe Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of task evaluations, the 3.5 series consistently demonstrates performance on par with state-of-the-art leading models. Compared to the 3 series, these models show a leap forward in both pure-text and multimodal capabilities.](/ai-gateway/models/qwen3.5-plus)
- [Alibaba CloudQwen 3.6 27BThe Qwen3.6 35B-A3B native vision-language model is built on a hybrid architecture that integrates linear attention mechanisms with a sparse mixture-of-experts framework, achieving higher inference efficiency. Compared with the 3.5-35B-A3B, this model demonstrates significantly improved agentic coding capabilities, mathematical and code reasoning abilities, spatial intelligence, as well as object localization and object detection performance.](/ai-gateway/models/qwen3.6-27b)
- [Alibaba CloudQwen 3.6 Max PreviewCompared with the previously released Qwen3-Max and Qwen3.6-Plus, this model features enhanced vibe coding abilities, more efficient coding agent execution, and significantly improved front-end development skills. Additionally, its long-tail knowledge retention has been further upgraded.](/ai-gateway/models/qwen-3.6-max-preview)
- [Alibaba CloudQwen 3.6 PlusThe Qwen3.6 native vision-language Plus series models demonstrate exceptional performance on par with the current state-of-the-art models, with a significant improvement in overall results compared to the 3.5 series. The models have been markedly enhanced in code-related capabilities such as agentic coding, front-end programming, and Vibe coding, as well as in multi-modal general object recognition, OCR, and object localization.](/ai-gateway/models/qwen3.6-plus)
- [Alibaba CloudQwen 3.7 FlashThe Qwen3.7 native vision-language Flash model series delivers a comprehensive upgrade over 3.6-Flash in multimodal understanding and agent execution. This model particularly excels in enhanced multimodal foundations with stronger universal object recognition, further improved real-world perception and spatial intelligence, significantly upgraded multimodal agent capabilities for Search Agent and CI Agent scenarios with more stable end-to-end task execution, as well as optimized multimodal coding for a smoother vibe coding experience.](/ai-gateway/models/qwen3.7-flash)
- [Alibaba CloudQwen 3.7 MaxQwen3.7 is a next‑generation flagship model designed for the agent‑centric era, with its core strengths lying in the breadth and depth of its agent‑level capabilities: it excels at programming, office and productivity tasks, and long‑term autonomous execution.](/ai-gateway/models/qwen3.7-max)
- [Alibaba CloudQwen 3.7 PlusAmong the Qwen3.7 series, the cost-effective Plus model builds on its robust text capabilities while delivering a comprehensive upgrade to its vision‑language abilities, all while preserving its full‑stack agent‑level intelligence for coding, tool use, and productivity workflows.](/ai-gateway/models/qwen3.7-plus)
- [Alibaba CloudQwen 3.8 FlashQwen3.8-Flash is Qwen’s fast, cost-efficient multimodal model, combining advanced reasoning and generation with a native 1M-token context window. Built for coding, agentic workflows, and visual understanding, it handles large codebases, long documents, charts, videos, and desktop applications.](/ai-gateway/models/qwen3.8-flash)
- [Alibaba CloudQwen 3.8 Flash NextQwen3.8-Flash-Next is Qwen’s experimental open-weight multimodal language model, pairing 125B parameters with just 6B activated for efficient reasoning and generation. With native 262K-token context extensible to 1M, vision support, and an architecture optimized for lower long-context latency, it is built for demanding agentic coding, tool use, visual reasoning, and multimodal automation.](/ai-gateway/models/qwen3.8-flash-next)
- [Alibaba CloudQwen 3.8 MaxQwen 3.8 Max is a 2.4-trillion-parameter MoE model delivering a comprehensive leap in coding and professional work. Autonomously codes and delivers complete projects spanning 10+ days. Handles hundreds of specialized tasks across legal, financial, design, and other professional domains, producing production-grade results end-to-end in a single conversation. Native visual understanding runs through the full cycle of planning, execution, and verification, enabling deep semantic analysis of ultra-long documents and extended video content. In long-horizon tasks, plans autonomously, iterates through closed feedback loops, and continuously evolves.](/ai-gateway/models/qwen3.8-max)
- [Alibaba CloudQwen3 235B A22BQwen3-235B-A22B-Instruct-2507 is the updated version of the Qwen3-235B-A22B non-thinking mode, featuring Significant improvements in general capabilities, including instruction following, logical reasoning, text comprehension, mathematics, science, coding and tool usage.](/ai-gateway/models/qwen-3-235b)
- [Alibaba CloudQwen3 235B A22B Thinking 2507Qwen3-235B-A22B-Thinking-2507 is the Qwen3's new model with scaling the thinking capability of Qwen3-235B-A22B, improving both the quality and depth of reasoning.](/ai-gateway/models/qwen3-235b-a22b-thinking)
- [Alibaba CloudQwen3 Coder 480B A35B InstructQwen3-Coder-480B-A35B-Instruct is Qwen's most agentic code model, featuring significant performance on Agentic Coding, Agentic Browser-Use and other foundational coding tasks, achieving results comparable to Claude Sonnet.](/ai-gateway/models/qwen3-coder)
- [Alibaba CloudQwen3 Coder NextQwen3-Coder-Next is an open-weight language model built specifically for coding, with strong performance on large-scale software engineering and agentic coding benchmarks. It uses a hybrid Mixture-of-Experts architecture to offer high capability at relatively modest active parameter counts, improving efficiency for real-world deployments. The model is trained on diverse code and natural language data so it can handle tasks like code generation, refactoring, debugging, repository-level reasoning, and technical explanation across multiple programming languages. It is also optimized for tool use and function calling, making it suitable as the core of coding agents that interact with shells, editors, issue trackers, and other developer tools.](/ai-gateway/models/qwen3-coder-next)
- [Alibaba CloudQwen3 Coder PlusPowered by Qwen3 this is a powerful Coding Agent that excels in tool calling and environment interaction to achieve autonomous programming. It combines outstanding coding proficiency with versatile general-purpose abilities.](/ai-gateway/models/qwen3-coder-plus)
- [Alibaba CloudQwen3 Embedding 0.6BThe Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks. Building upon the dense foundational models of the Qwen3 series, it provides a comprehensive range of text embeddings and reranking models in various sizes (0.6B, 4B, and 8B).](/ai-gateway/models/qwen3-embedding-0.6b)
- [Alibaba CloudQwen3 Embedding 4BThe Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks. Building upon the dense foundational models of the Qwen3 series, it provides a comprehensive range of text embeddings and reranking models in various sizes (0.6B, 4B, and 8B).](/ai-gateway/models/qwen3-embedding-4b)
- [Alibaba CloudQwen3 Embedding 8BThe Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks. Building upon the dense foundational models of the Qwen3 series, it provides a comprehensive range of text embeddings and reranking models in various sizes (0.6B, 4B, and 8B).](/ai-gateway/models/qwen3-embedding-8b)
- [Alibaba CloudQwen3 MaxThe Qwen 3 series Max model has undergone specialized upgrades in agent programming and tool invocation compared to the preview version. The officially released model this time has achieved state-of-the-art (SOTA) performance in its field and is better suited to meet the demands of agents operating in more complex scenarios.](/ai-gateway/models/qwen3-max)
- [Alibaba CloudQwen3 Max PreviewQwen3-Max-Preview shows substantial gains over the 2.5 series in overall capability, with significant enhancements in Chinese-English text understanding, complex instruction following, handling of subjective open-ended tasks, multilingual ability, and tool invocation; model knowledge hallucinations are reduced.](/ai-gateway/models/qwen3-max-preview)
- [Alibaba CloudQwen3 Next 80B A3B InstructA new generation of open-source, non-thinking mode model powered by Qwen3. This version demonstrates superior Chinese text understanding, augmented logical reasoning, and enhanced capabilities in text generation tasks over the previous iteration (Qwen3-235B-A22B-Instruct-2507).](/ai-gateway/models/qwen3-next-80b-a3b-instruct)
- [Alibaba CloudQwen3 Next 80B A3B ThinkingA new generation of Qwen3-based open-source thinking mode models. This version offers improved instruction following and streamlined summary responses over the previous iteration (Qwen3-235B-A22B-Thinking-2507).](/ai-gateway/models/qwen3-next-80b-a3b-thinking)
- [Alibaba CloudQwen3 VL 235B A22B InstructThe Qwen3 series VL models has been comprehensively upgraded in areas such as visual coding and spatial perception. Its visual perception and recognition capabilities have significantly improved, supporting the understanding of ultra-long videos, and its OCR functionality has undergone a major enhancement.](/ai-gateway/models/qwen3-vl-instruct)
- [Alibaba CloudQwen3 VL 235B A22B ThinkingQwen3 series VL models feature significantly enhanced multimodal reasoning capabilities, with a particular focus on optimizing the model for STEM and mathematical reasoning. Visual perception and recognition abilities have been comprehensively improved, and OCR capabilities have undergone a major upgrade.](/ai-gateway/models/qwen3-vl-thinking)
- [Alibaba CloudQwen3-14BQwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities, and multilingual support](/ai-gateway/models/qwen-3-14b)
- [Alibaba CloudQwen3-30B-A3BQwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities, and multilingual support](/ai-gateway/models/qwen-3-30b)
- [Alibaba CloudQwen3.8 2.4T A95BOpen-weights release of the Qwen3.8 flagship (2.4T MoE, ~95B active); thinking always on with reasoning\_effort low/medium/xhigh. The hosted Qwen 3.8 Max (vision, non-thinking, 1M default context) is Alibaba-only.](/ai-gateway/models/qwen3.8-2.4t-a95b)
- [Alibaba CloudQwen3.8 27BBuilt on the architectural foundation of Qwen3.5, Qwen3.8 delivers substantial gains across coding, professional work, research, and long-horizon agentic tasks. Qwen3.8-27B brings these advances to a compact, deployment-friendly dense model: a native vision-language model that understands images and videos, with flexible thinking control, designed to carry complex, multi-step tasks through to completion with greater reliability.](/ai-gateway/models/qwen3.8-27b)
- [RecraftRecraft V2Recraft V2 is an image generation model released in March 2024 and the first model trained from scratch by Recraft. With 20 billion parameters, it was a breakthrough in human anatomical accuracy and the first to support brand consistency and brand color inputs. It also introduced vector image generation (SVG output), as well as minimalistic icon and illustration styles.](/ai-gateway/models/recraft-v2)
- [RecraftRecraft V3V3 introduced major advances in photorealism and text rendering. It was the first Recraft model to generate mid-size text accurately and, as of 2025, is the only model capable of placing text at specific positions in an image.](/ai-gateway/models/recraft-v3)
- [RecraftRecraft V4The model delivers strong photorealism, including realistic skin rendering and natural textures, while avoiding common synthetic artifacts. It produces more distinctive lighting, composition, diverse subjects, contemporary styling, and carefully considered scene elements. For illustration, it generates original characters and forms with sophisticated and unexpected color combinations.](/ai-gateway/models/recraft-v4)
- [RecraftRecraft V4 ProThe model delivers strong photorealism, including realistic skin rendering and natural textures, while avoiding common synthetic artifacts. It produces more distinctive lighting, composition, diverse subjects, contemporary styling, and carefully considered scene elements. For illustration, it generates original characters and forms with sophisticated and unexpected color combinations.](/ai-gateway/models/recraft-v4-pro)
- [RecraftRecraft V4.1V4.1 is built on the same visual aesthetic and the same eye for what looks right, just pushed further across every dimension. The photorealism is more natural. The gradients are dreamier. Illustration styles are now possible that simply weren't before. And this is a model that reads a short prompts easier and creates something worth stopping for.](/ai-gateway/models/recraft-v4.1)
- [RecraftRecraft V4.1 ProV4.1 is built on the same visual aesthetic and the same eye for what looks right, just pushed further across every dimension. The photorealism is more natural. The gradients are dreamier. Illustration styles are now possible that simply weren't before. And this is a model that reads a short prompts easier and creates something worth stopping for. V4.1 Pro generates higher-resolution images for when the idea deserves more room.](/ai-gateway/models/recraft-v4.1-pro)
- [RecraftRecraft V4.1 UtilityV4.1 is built on the same visual aesthetic and the same eye for what looks right, just pushed further across every dimension. The photorealism is more natural. The gradients are dreamier. Illustration styles are now possible that simply weren't before. And this is a model that reads a short prompts easier and creates something worth stopping for. V4.1 Utility is designed for when restraint is the aesthetic choice, with flat lighting, front-facing composition, and simple, controlled scenes.](/ai-gateway/models/recraft-v4.1-utility)
- [RecraftRecraft V4.1 Utility ProV4.1 is built on the same visual aesthetic and the same eye for what looks right, just pushed further across every dimension. The photorealism is more natural. The gradients are dreamier. Illustration styles are now possible that simply weren't before. And this is a model that reads a short prompts easier and creates something worth stopping for. V4.1 Utility is designed for when restraint is the aesthetic choice, with flat lighting, front-facing composition, and simple, controlled scenes.](/ai-gateway/models/recraft-v4.1-utility-pro)
- [Fish AudioS1Fish Audio S1 is trained on over 2 million hours of audio with online RLHF (GRPO). It achieves 0.8% WER and 0.4% CER on Seed TTS Eval. S1 supports open-domain emotion, tone, and special effect markers.](/ai-gateway/models/s1)
- [Fish AudioS2 ProFish Audio S2 is the second-generation TTS model from Fish Audio. It's trained on over 10 million hours of audio across approximately 80 languages, and it introduces inline tag control: natural-language instructions embedded directly in your script at any position, giving you fine-grained direction over how speech is delivered at the word or phrase level.](/ai-gateway/models/s2-pro)
- [Fish AudioS2.1 ProS2.1 Pro is Fish Audio's current state-of-the-art voice model — the best model we have, now available to every developer for free via API. It is a neural speech synthesis model designed for production-grade AI voice generation, with particular strengths in low-latency streaming, multilingual TTS, and voice cloning.](/ai-gateway/models/s2.1-pro)
- [Sakana AISakana NamazuSakana Namazu is a Japanese-specialized LLM that combines a deep understanding of Japanese culture and business customs with high-performance language capabilities. Built on the open model Kimi K2.6 and refined with Sakana AI's in-house data for Japanese language and business workflows, it handles complex tasks using web search and code execution. Unlike Fugu, which orchestrates multiple frontier models, Sakana Namazu provides a single in-house model as an API.](/ai-gateway/models/namazu)
- [ByteDanceSeed 1.6ByteDance's new multimodal deep-thinking model, supporting both text and visual inputs with enhanced reasoning capabilities.](/ai-gateway/models/seed-1.6)
- [ByteDanceSeedance 2.0Built with a unified multimodal audio-video joint generation architecture, Seedance 2.0 supports four input modalities: text, image, audio, and video. Compared with Version 1.5, Seedance 2.0 delivers a substantial leap in generation quality. It achieves a higher usability rate for complex interaction and motion scenes, with significant improvements in physical accuracy, visual realism, and controllability, making it well-suited for high-quality creation scenarios.](/ai-gateway/models/seedance-2.0)
- [ByteDanceSeedance 2.0 FastSeedance 2.0 Fast is a new-generation multimodal video creation model, inheriting the core functions and advantages of Seedance 2.0, with faster speed.](/ai-gateway/models/seedance-2.0-fast)
- [ByteDanceSeedance 2.0 MiniSeedance 2.0 Mini is a cost-effective multimodal video creation model, retaining the core functions and advantages of Seedance 2.0 at a lower price.](/ai-gateway/models/seedance-2.0-mini)
- [ByteDanceSeedance 2.5Seedance 2.5 is a next-generation audio-video joint generation model, built for 30-second storytelling with precise reference control and powerful editing capabilities.](/ai-gateway/models/seedance-2.5)
- [ByteDanceSeedance v1.0 ProA video generation model that supports multi-shot storytelling. It excels in semantic understanding and instruction following, producing smooth, detailed, and cinematic 1080P HD videos.](/ai-gateway/models/seedance-v1.0-pro)
- [ByteDanceSeedance v1.0 Pro FastSeedance 1.0 Pro Fast delivers top performance at an unbeatable price, balancing quality, speed, and cost. Built on Seedance 1.0 Pro’s core strengths, it’s faster and more cost-efficient for creators.](/ai-gateway/models/seedance-v1.0-pro-fast)
- [ByteDanceSeedance v1.5 ProByteDance's Seedance 1.5 Pro is a professional video model using V2A native generation for integrated, synced audio-visual output, enhancing efficiency of professional video creation.](/ai-gateway/models/seedance-v1.5-pro)
- [ByteDanceSeedream 4.0Seedream 4.0 is a SOTA multimodal image creation model built on leading architecture. It breaks through the boundaries of traditional text-to-image models by natively supporting text, single-image, and multi-image inputs. Users can freely combine text and images to achieve diverse creative modes within a single model—such as multi-image blending, image editing, and sequentially batch image generation, featuring subject consistency, making image creation more free and controllable.](/ai-gateway/models/seedream-4.0)
- [ByteDanceSeedream 4.5Seedream 4.5 is the latest in-house image generation model developed by ByteDance. Compared with Seedream 4.0, it delivers comprehensive improvements—especially in editing consistency, including better preservation of subject details, lighting, and color tone. It also enhances portrait refinement and small-text rendering. The model’s multi-image composition capabilities have been significantly strengthened, and both reasoning performance and visual aesthetics continue to advance, enabling more accurate and artistically expressive image generation.](/ai-gateway/models/seedream-4.5)
- [ByteDanceSeedream 5.0 LiteByteDance-Seedream-5.0-lite is the latest image generation model released by BytePlus. For the first time, it introduces web-connected retrieval, enabling the model to fuse real-time online information to significantly improve the timeliness and relevance of generated images. The model’s reasoning and comprehension capabilities are further upgraded, allowing it to accurately interpret complex prompts and visual inputs. In addition, ByteDance-Seedream-5.0-lite delivers notable improvements in global knowledge coverage, reference consistency, and professional-grade scene generation, making it well suited for enterprise-level visual creation workflows.](/ai-gateway/models/seedream-5.0-lite)
- [ByteDanceSeedream 5.0 ProSeedream-5.0-Pro, ByteDance's newest image generation model, delivers comprehensive upgrades for complex, lifelike image creation and editing, ushering in a new phase of controllable visual production. It stands out with precise editing control, robust commercial applicability and natural rendering results.](/ai-gateway/models/seedream-5.0-pro)
- [PerplexitySonarPerplexity's lightweight offering with search grounding, quicker and cheaper than Sonar Pro.](/ai-gateway/models/sonar)
- [PerplexitySonar ProPerplexity's premier offering with search grounding, supporting advanced queries and follow-ups.](/ai-gateway/models/sonar-pro)
- [PerplexitySonar Reasoning ProA premium reasoning-focused model that outputs Chain of Thought (CoT) in responses, providing comprehensive explanations with enhanced search capabilities and multiple search queries per request.](/ai-gateway/models/sonar-reasoning-pro)
- [StepFunStep 3.7 FlashStepFun’s flagship multimodal reasoning model. Powered by a 198B-parameter / 11B-activation sparse MoE architecture, with native support for image and video understanding.](/ai-gateway/models/step-3.7-flash)
- [StepFunStepFun 3.5 FlashStep 3.5 Flash is an open-source reasoning model by StepFun with 196B total parameters (11B active) using Mixture of Experts. It features a 256K context window, deep reasoning, tool calling, and agentic capabilities, achieving 97.3 on AIME 2025 and 74.4% on SWE-bench Verified.](/ai-gateway/models/step-3.5-flash)
- [TakoTako SearchSearch Tako's curated knowledge graph and the live web for source-grounded data cards, citations, and visualizations. Use it as a built-in tool through the AI Gateway to give any model access to structured data and current web information.](/ai-gateway/models/search)
- [Tencent CloudTencent Hy-MT2-LiteHy-MT2-Lite is Tencent Cloud Hunyuan’s lightweight 1.8B translation model, combining efficient performance with an 8K-token context window and broad multilingual coverage. It supports translation across 33 languages and five ethnic Chinese and dialect variants, delivering strong results on benchmarks including FLORES-200 and WMT25. Designed for professional and real-world business use, it reliably follows instructions for structured, delimiter-preserving, context-aware, glossary-guided, and style-specific translation.](/ai-gateway/models/hy-mt2-lite)
- [Tencent CloudTencent Hy-MT2-PlusHy-MT2-Plus is Tencent Cloud Hunyuan’s 7B translation model, offering an 8K-token context window and broad multilingual coverage. Optimized for translation across 33 languages and five ethnic Chinese and dialect variants, it delivers industry-leading performance on benchmarks including FLORES-200 and WMT25. It excels in professional domains and real-world business scenarios, with strong instruction-following for structured, delimiter-preserving, context-aware, glossary-guided, and style-specific translation.](/ai-gateway/models/hy-mt2-plus)