Skip to content
Dashboard

Qwen3 Next 80B A3B Instruct

Qwen3 Next 80B A3B Instruct is an 80-billion-parameter hybrid Transformer-Mamba model that activates only 3B parameters per token, delivering high inference throughput over dense alternatives at a native context window of 262.1K tokens.

Input and output price
Prices from: Input $0.09, Output $1.10, Per 1M tokens
24h uptime
Loading AI Gateway uptime
import { streamText } from 'ai'
const result = streamText({
model: 'alibaba/qwen3-next-80b-a3b-instruct',
prompt: 'Why is the sky blue?'
})
Read docs

Copy link to headingMore models by Alibaba Cloud

Model
Context
Latency
Throughput
Input
Output
Cache
Web Search
Capabilities
Providers
ZDR
No Training
Free Tier
Release Date
991K
3.3 s
57 tps
$2/M
$6/M
Read:$0.25/M
Write:$2.50/M
+2
alibaba logo
09/01/2026
991K
3.5 s
86 tps
$0.16/M
$0.47/M
Read:$0.02/M
Write:$0.20/M
+2
alibaba logo
08/26/2026
1M
0.4 s
129 tps
$0.10/M
$0.40/M
Read:$0.01/M
Write:$0.63/M
+1
alibaba logo
deepinfra logo
morph logo
+2
08/14/2026
1M
3.9 s
88 tps
$2/M
$6/M
Read:$0.25/M
Write:$2.50/M
+1
alibaba logo
fireworks logo
08/02/2026
991K
1.3 s
142 tps
$0.03/M+2 more
$0.13/M+2 more
Read:
$0.006/M+2 more
Write:
$0.04/M+2 more
+2
alibaba logo
07/28/2026
1M
1.9 s
58 tps
$0.40/M+1 more
$1.60/M+1 more
Read:
$0.08/M+1 more
Write:
$0.50/M+1 more
+2
alibaba logo
06/02/2026

Your use is subject to Alibaba Cloud's Terms & Privacy Policies.