Minimax
100.00%Context205K
Max output205K
p50 latency946 ms
Throughput70.0 t/s
Input price$0.3
Output price$1.2
minimax/minimax-m2
MiniMax-M2 redefines efficiency for agents. It is a compact, fast, and cost-effective MoE model (230 billion total parameters with 10 billion active parameters) built for elite performance in coding and agentic tasks, all while maintaining powerful general intelligence.
USD per one million tokens, using current provider data.
Last-hour p50 in milliseconds; lower is better.
Reported availability over the last 24 hours; not a contractual SLA.
Compare current pricing, performance, and privacy across every available provider.
| Provider | Maximum context | Max output | p50 latency | p50 throughput | 24h uptime | Input / 1M tokens | Output / 1M tokens | Privacy |
|---|---|---|---|---|---|---|---|---|
Minimax | 205K | 205K | 946 ms | 70.0 t/s | 100.000% | $0.3 | $1.2 | |
Novita | 204.8K | 131.1K | 1,284 ms | 76.5 t/s | 100.000% | $0.3 | $1.2 |
Available modalities and capabilities across inference providers.
minimax/minimax-m2More options from the same creator or model category.