Skip to content
Dashboard

MiniMax M2.1

MiniMax M2.1 is MiniMax's second-generation model, focused on coding accuracy, tool use, instruction following, and long-horizon planning. It supports a context window of 204.8K tokens and a max output of 131.1K tokens per request. Your use is subject to MiniMax's Terms & Privacy Policies.

ReasoningTool UseImplicit Caching
import { streamText } from 'ai'
const result = streamText({
model: 'minimax/minimax-m2.1',
prompt: 'Why is the sky blue?'
})
Read docs

Copy link to headingMore models by MiniMax

Model
Context
Latency
Throughput
Input
Output
Cache
Web Search
Capabilities
Providers
ZDR
No Training
Free Tier
Release Date
1M
0.6s
142tps
Free
Free
Free
+2
fireworks logo
gmicloud logo
minimax logo
+2
05/31/2026
205K
0.8s
92tps
Free
Free
Free
gmicloud logo
minimax logo
novita logo
03/18/2026
205K
1.0s
47tps
$0.60/M
$2.40/M
Read:$0.06/M
Write:$0.38/M
minimax logo
03/18/2026
1M
0.4s
103tps
$0.27/MFast $0.60/M
$0.95/MFast $2.40/M
Read:$0.03/M
Write:$0.38/M
bedrock logo
deepinfra logo
minimax logo
+1
02/12/2026
205K
1.1s
67tps
$0.60/M
$2.40/M
Read:$0.03/M
Write:$0.38/M
minimax logo
novita logo
02/12/2026
205K
0.8s
81tps
$0.30/M
$1.20/M
Read:$0.03/M
Write:$0.38/M
minimax logo
novita logo
10/27/2025