Last refreshed 2026-06-15. Next refresh: weekly.
Why use MiniMax M2.5 Highspeed on MiniMax?
MiniMax offers MiniMax M2.5 Highspeed with competitive pricing. MiniMax is a multimodal foundation model and API platform for text, speech, video, image, and music generation with agent tools.
Compare MiniMax M2.5 Highspeed across 3 providers to find the best fit for your use caseInput / 1M
-
Output / 1M
-
Cache
Not sourced
Batch
Not sourced
Setup recipe
Docs fallbackInstall
Use the provider REST API or SDKAuth
Create a provider API keyCall
model: MiniMax-M2.5-highspeedModel ID
MiniMax-M2.5-highspeedRequest example
Curated snippets for this provider are not sourced yet. Use MiniMax documentation with model ID
MiniMax-M2.5-highspeed.Gotchas
- Use provider model ID "MiniMax-M2.5-highspeed", not the LLMReference slug "minimax-m2.5-highspeed".
Compare MiniMax M2.5 Highspeed Across Providers
| Provider | Input (per 1M) | Output (per 1M) |
|---|---|---|
| MiniMax | — | — |
| Vercel AI Gateway | $0.60 | $2.40 |
| Novita AI | $0.60 | $2.40 |
Capabilities
ReasoningJSON / Tool useStructured Outputs
About MiniMax M2.5 Highspeed
MiniMax M2.5 Highspeed is MiniMax's inference-optimized variant of M2.5, released simultaneously in February 2026. It delivers identical intelligence and outputs to standard M2.5 through a specialized inference engine at lower latency. The model supports a 204,800-token context window, 131,072-token max output, function calling, structured output, and reasoning. API model ID: MiniMax-M2.5-highspeed. It is designed for latency-sensitive interactive applications and automated agent pipelines.
Get Started
Model Specs
Released2026-02-12
Parameters230B (10B active)
Context205k
ArchitectureDecoder Only