MiniMax-M2 redefines efficiency for intelligent agents. It is a compact, fast, and cost-effective MoE model with a total of 230 billion parameters and 10 billion active parameters, designed for top performance in coding and intelligent agent tasks while maintaining strong general intelligence. With only 10 billion active parameters, MiniMax-M2 delivers the complex end-to-end tool usage performance expected from today's leading models, but in a more streamlined form factor, making deployment and scaling easier than ever before.
Pricing
- Input Tokens: $0.288 /M tokens
- Output Tokens: $1.152 /M tokens
Input Modalities
- Text
Output Modalities
- Text
Context length
- 205K tokens
Max output
- 205K tokens
Capabilities
- Thinking
- Streaming
- Tool calling
- Web search
- URL context
- Code interpreter
- Computer use
- File search
- Memory tool
- Structured outputs
- Citations
- Prompt caching
- Background mode
- Server-side sessions
Providers
Minimax minimax-m2
Pricing$0.288$1.152
Context204K
Max output131K
Latency3.1S
Throughput64.8TPS
Uptime
100.00% uptime 3 days ago
100.00% uptime 2 days ago
0.00% uptime yesterday
Siliconflow minimax-m2
Pricing$0.288$1.152
Context192K
Max output192K
Latency1.4S
Throughput16.7TPS
Uptime
0.00% uptime 3 days ago
0.00% uptime 2 days ago
0.00% uptime yesterday
Performance for minimax-m2
Uptime is the percentage of requests that succeeded over the past 72 hours. AIHubMix continuously monitors every provider and automatically retries with the next-best provider when one returns an error or responds too slowly; Latency is total round-trip time (lower is better); Throughput is how fast the model writes (tokens per second, higher is better).
Uptime
Loading...
Latency
Loading...
Throughput
Loading...
Try this model
Python