MiniMax-M2.5 is a SOTA large language model designed for real-world productivity. Trained in a diverse range of complex real-world digital working environments, M2.5 builds upon the coding expertise of M2.1 to extend into general office work, reaching fluency in generating and operating Word, Excel, and Powerpoint files, context switching between diverse software environments, and working across different agent and human teams. Scoring 80.2% on SWE-Bench Verified, 51.3% on Multi-SWE-Bench, and 76.3% on BrowseComp, M2.5 is also more token efficient than previous generations, having been trained to optimize its actions and output through planning.
- By:
- MiniMax
- Input:
- Output:
- Context length:
- 205K
- Max output:
- 131K
- Published:
- 2026-02-12
Providers
Same model, different providers. Automatic routing picks the cheapest healthy one within the same quality tier.
| Provider | ||||
|---|---|---|---|---|
| 20% offminimax | $0.3$0.24/1M tokens | $1.2$0.96/1M tokens | Read$0.03$0.024/1M tokensWrite$0.375$0.3/1M tokens | 205K |
Availability
24 hoursSuccess rate of the requests Amux actually sent to each provider. Client errors are excluded from the denominator — a malformed request isn't the provider's fault.
Throughput
24 hoursAverage output speed on streaming requests, in tokens per second (TPS).
Latency
24 hoursAverage time to first token (TTFT) on streaming requests — how long from sending a request to receiving its first token.
Activity
24 hoursToken usage and cost for this model over time, split by the provider that served each request.
Related models
More models from MiniMax