[Alibaba Cloud](/ai-gateway/models/labs/alibaba)

# Qwen3 235B A22B

Qwen3 235B A22B is Alibaba Cloud's large-scale 235B mixture-of-experts model with a context window of 262.1K tokens, activating 22 billion of 235 billion parameters per inference to deliver strong reasoning, coding, and multilingual performance. Your use is subject to Alibaba Cloud's [Terms](https://www.alibabacloud.com/help/en/legal/latest/alibaba-cloud-international-website-product-terms-of-service-v-3-8-0) & [Privacy](https://www.alibabacloud.com/help/en/legal/latest/alibaba-cloud-international-website-privacy-policy) Policies.

Tool Use

[Use with AI Gateway](https://vercel.com/d?to=%2F%5Bteam%5D%2F%7E%2Fai%3Futm_source%3Dgateway-model-page%26utm_campaign%3Dai-gateway-models&title=Get+Started+with+Vercel+AI+Gateway) [View docs](https://vercel.com/docs/ai-gateway)

AI SDKChat CompletionsMessagesResponses

```
1import { streamText } from 'ai'
2

3const result = streamText({
4  model: 'alibaba/qwen-3-235b',
5  prompt: 'Why is the sky blue?'
6})
```

[Read docs](https://vercel.com/docs/ai-gateway/sdks-and-apis/ai-sdk)

[Overview](/ai-gateway/models/qwen-3-235b) [API](/ai-gateway/models/qwen-3-235b/api) [About](/ai-gateway/models/qwen-3-235b/about) [Providers](/ai-gateway/models/qwen-3-235b/providers) [Throughput](/ai-gateway/models/qwen-3-235b/throughput) [Latency](/ai-gateway/models/qwen-3-235b/latency) [Uptime](/ai-gateway/models/qwen-3-235b/uptime) [Status](/ai-gateway/models/qwen-3-235b/status) [Similar](/ai-gateway/models/qwen-3-235b/similar) [FAQ](/ai-gateway/models/qwen-3-235b/faq)

## [Copy link to heading](#playground)Playground

Try out Qwen3 235B A22B by Alibaba Cloud. Usage is billed to your team at API rates. Free users (those who haven't made a payment) get $5 of credits every 30 days.

![alibaba logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Falibaba%2520cloud.png&w=48&q=75)Qwen3 235B A22B

![alibaba logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Falibaba%2520cloud.png&w=96&q=75)

Qwen3 235B A22B

## [Copy link to heading](#providers)Providers

Route requests across multiple providers. Copy a provider slug to set your preference. Visit the [docs](/docs/ai-gateway/provider-options) for more info. Using a provider means you agree to their terms, listed under Legal.

| Provider |
| --- |

| Context | Max Output | Latency | Throughput | Input | Output | Cache | Web Search | Capabilities | ZDR | No Training | Release Date |  |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- |

| ![deepinfra logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepinfra.png&w=48&q=75) [DeepInfra](/ai-gateway/models/providers/deepinfra) Legal:[Terms](https://deepinfra.com/terms)•[Privacy](https://deepinfra.com/privacy) | 41K | 16K | 0.4s | 13tps | $0.09/M | $0.10/M |  | — |  |  |  | 04/28/2025 |  |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- |
| ![vertex logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fvertex%2520ai.png&w=48&q=75) [Google Vertex AI](/ai-gateway/models/providers/vertex) Legal:[Terms](https://cloud.google.com/terms/service-terms)•[Privacy](https://cloud.google.com/privacy) | 262K | 16K | 0.5s | 20tps | $0.22/M | $0.88/M |  | — |  |  |  | 04/28/2025 |  |
| ![novita logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fnovita.png&w=48&q=75) [Novita AI](/ai-gateway/models/providers/novita) Legal:[Terms](https://novita.ai/legal/terms-of-service)•[Privacy](https://novita.ai/legal/privacy-policy) | 131K | 33K | 0.6s | 73tps | $0.09/M | $0.58/M |  | — |  |  |  | 04/28/2025 |  |

## [Copy link to heading](#throughput)Throughput24 hours

1W

1D

P50 throughput on live AI Gateway traffic, in tokens per second (TPS). Visit the [docs](https://vercel.com/docs/ai-gateway/models-and-providers/metrics) for more info.

## [Copy link to heading](#latency)Latency24 hours

1W

1D

P50 time to first token (TTFT) on live AI Gateway traffic, in milliseconds. View the [docs](https://vercel.com/docs/ai-gateway/models-and-providers/metrics) for more info.

## [Copy link to heading](#uptime)Uptime24 hours

1W

1D

1H

Direct request success rate on AI Gateway and per-provider. Visit the [docs](https://vercel.com/docs/ai-gateway/models-and-providers/uptime) for more info.

1W

1D

1H

## [Copy link to heading](#more-models-by-alibaba-cloud)More models by Alibaba Cloud

All

Text

Code

| Model |
| --- |

| Context | Latency | Throughput | Input | Output | Cache | Web Search | Capabilities | Providers | ZDR | No Training | Release Date |  |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- |

| ![alibaba logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Falibaba%20model%20logo.png&w=48&q=75) [alibaba/qwen3.8-27b](/ai-gateway/models/qwen3.8-27b) | 1M | 0.3s | 123tps | $0.10/M | $0.40/M | Read:$0.01/M Write:— | — | +1 | ![alibaba logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Falibaba%2520cloud.png&w=48&q=75) ![deepinfra logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepinfra.png&w=48&q=75) ![runinfra logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fruninfra.png%3Fv%3D1786910657446&w=48&q=75) |  |  | 08/14/2026 |  |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- |
| ![alibaba logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Falibaba%20model%20logo.png&w=48&q=75) [alibaba/qwen3.8-2.4t-a95b](/ai-gateway/models/qwen3.8-2.4t-a95b) | 262K | 0.4s | 72tps | $1.65/M | $4.95/M | Read:$0.12/M Write:$2.5/M | — |  | ![deepinfra logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepinfra.png&w=48&q=75) ![fireworks logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Ffireworks.png&w=48&q=75) ![gmicloud logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fgmicloud.png%3Fv%3D1785888489595&w=48&q=75) +3 |  |  | 08/03/2026 |  |
| ![alibaba logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Falibaba%20model%20logo.png&w=48&q=75) [alibaba/qwen3.8-max](/ai-gateway/models/qwen3.8-max) | 1M | 4.1s | 122tps | $2/M | $6/M | Read:$0.25/M Write:$2.5/M | — | +1 | ![alibaba logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Falibaba%2520cloud.png&w=48&q=75) ![fireworks logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Ffireworks.png&w=48&q=75) |  |  | 08/02/2026 |  |
| ![alibaba logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Falibaba%20model%20logo.png&w=48&q=75) [alibaba/qwen3.7-flash](/ai-gateway/models/qwen3.7-flash) | 991K | 3.2s | 132tps | $0.03/M+2 more | $0.13/M+2 more | Read: $0.01/M+2 more Write: $0.04/M+2 more | — | +2 | ![alibaba logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Falibaba%2520cloud.png&w=48&q=75) |  |  | 07/28/2026 |  |
| ![alibaba logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Falibaba%20model%20logo.png&w=48&q=75) [alibaba/qwen3.7-plus](/ai-gateway/models/qwen3.7-plus) | 1M | 1.3s | 230tps | $0.40/M+1 more | $1.60/M+1 more | Read: $0.08/M+1 more Write: $0.5/M+1 more | — | +2 | ![alibaba logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Falibaba%2520cloud.png&w=48&q=75) ![fireworks logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Ffireworks.png&w=48&q=75) |  |  | 06/02/2026 |  |
| ![alibaba logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Falibaba%20model%20logo.png&w=48&q=75) [alibaba/qwen3-vl-235b-a22b-instruct](/ai-gateway/models/qwen3-vl-235b-a22b-instruct) | 262K | 0.8s | 52tps | $0.20/M | $0.88/M | Read:$0.11/M Write:— | — |  | ![alibaba logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Falibaba%2520cloud.png&w=48&q=75) ![deepinfra logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepinfra.png&w=48&q=75) |  |  | 09/23/2025 |  |

## [Copy link to heading](#about-qwen3-235b-a22b)About Qwen3 235B A22B

Qwen3 235B A22B is the largest MoE in the Qwen3 family. Its MoE design routes each token through 8 of 128 expert layers, activating only 22 billion of the full 235 billion parameters. This sparsity keeps serving costs proportional to the activated parameter count while the full parameter space retains the breadth of knowledge needed for performance competitive with large proprietary models.

The model covers 119 languages and dialects.

On benchmark evaluations, Qwen3 235B A22B achieves results competitive with other strong reasoning models across coding, mathematics, and general capability assessments. Like all Qwen3 models, it supports a hybrid reasoning system: thinking mode activates extended computation for step-by-step problem solving, while non-thinking mode produces immediate responses when latency matters more than deliberation. You can configure a thinking budget per request to tune the cost-quality tradeoff dynamically.

The model supports tool calling and Model Context Protocol (MCP), making it suitable for multi-step workflows where the model orchestrates external tools or APIs. Alibaba Cloud recommends pairing it with the Qwen-Agent framework for complex agentic pipelines.

## [Copy link to heading](#what-to-consider-when-choosing-a-provider)What To Consider When Choosing a Provider

- Configuration: For applications with strict data sovereignty requirements, provider selection through AI Gateway lets you target infrastructure in specific regions without modifying your application's API calls.
- Zero Data Retention: Zero Data Retention is available for this model. It is offered on a per-provider and model basis. See the [documentation](https://vercel.com/docs/ai-gateway/security-and-compliance/zdr) for details.
- Authentication: AI Gateway authenticates requests using an [API key](https://vercel.com/docs/ai-gateway/authentication-and-byok#api-key-authentication) or [OIDC token](https://vercel.com/docs/ai-gateway/authentication-and-byok#oidc-token-authentication). You do not need to manage provider credentials directly.

## [Copy link to heading](#when-to-use-qwen3-235b-a22b)When to Use Qwen3 235B A22B

### Best for

- Complex reasoning tasks: When a problem genuinely requires the highest reasoning ceiling in the Qwen3 family, such as graduate-level mathematics, complex logical deduction chains, or multi-step proofs, this model is designed for it
- Competitive coding challenges and complex software engineering: The model's benchmark results in coding tasks place it alongside large proprietary models, making it appropriate for problems that smaller models consistently struggle with
- Enterprise agentic pipelines: With strong tool-calling capabilities and MCP support, Qwen3 235B A22B can serve as the planning and reasoning backbone for multi-step automated workflows where reliability matters
- Multilingual content at scale: Covering 119 languages, this model suits global-facing products that need high-quality output across many language families without running separate per-language models
- Research and evaluation baselines: As the largest-capability open-weight model in the Qwen3 Apache 2.0-licensed line, it's a useful reference point for capability assessments

### Consider alternatives when

- Latency is a binding constraint: Larger models take longer to generate tokens. For interactive applications where response time outweighs peak quality, smaller Qwen3 variants or the 30B-A3B MoE model will serve users better
- Cost per query needs to be minimized: The 235B model carries higher inference costs than smaller family members. If most queries are routine instruction-following tasks, a smaller model will deliver adequate quality at lower cost
- Lighter-weight alternatives suffice: For teams that don't need the full capability ceiling of a 235B model, smaller dense variants like Qwen3-14B or Qwen3-32B handle many production workloads at lower per-token cost

## [Copy link to heading](#conclusion)Conclusion

Qwen3 235B A22B is the right choice when the task genuinely demands the highest capability in the Qwen3 family, complex reasoning, high-stakes coding, or multilingual generation at quality levels that smaller models can't reliably reach. Running it through AI Gateway means a single API integration covers provider failover, consolidated billing, and access to DeepInfra, Google Vertex AI, Novita AI without additional account management.