MiniMax-M2.5-Highspeed is a high-speed variant of MiniMax-M2.5, delivering the same SOTA performance and real-world productivity capabilities with significantly faster inference and greater responsiveness. Trained across diverse and complex digital working environments, it extends the coding expertise of M2.1 into general office work, including generating and operating Word, Excel, and PowerPoint files, switching seamlessly across software environments, and collaborating across agent and human teams. With the same strong benchmark performance and token-efficient planning capabilities as M2.5, the Highspeed variant is optimized for faster, more agile execution in latency-sensitive workflows.
- By:
- MiniMax
- Input:
- Output:
- Context length:
- 205K
- Max output:
- 131K
- Published:
- 2026-02-12
Providers
Same model, different providers. Automatic routing picks the cheapest healthy one within the same quality tier.
| Provider | ||||
|---|---|---|---|---|
| 20% offminimax | $0.6$0.48/1M tokens | $2.4$1.92/1M tokens | Read$0.03$0.024/1M tokensWrite$0.375$0.3/1M tokens | 205K |
Availability
24 hoursSuccess rate of the requests Amux actually sent to each provider. Client errors are excluded from the denominator — a malformed request isn't the provider's fault.
Throughput
24 hoursAverage output speed on streaming requests, in tokens per second (TPS).
Latency
24 hoursAverage time to first token (TTFT) on streaming requests — how long from sending a request to receiving its first token.
Activity
24 hoursToken usage and cost for this model over time, split by the provider that served each request.
Related models
More models from MiniMax