Last refreshed 2026-06-29. Next refresh: weekly.
Why use MiniMax M2.5 Highspeed on Novita AI?
Novita AI offers MiniMax M2.5 Highspeed with pay-as-you-go pricing at $0.60/1M input tokens. Novita AI offers a GPU-based inference API for image, video, and language model generation with a broad catalog of open-source models.
Compare MiniMax M2.5 Highspeed across 3 providers to find the best fit for your use caseSetup recipe
Docs fallbackUse the provider REST API or SDKCreate a provider API keymodel: minimax-m2.5-highspeedminimax-m2.5-highspeedRequest example
Gotchas
No curated gotchas have been sourced for this exact provider/model route yet.
Compare MiniMax M2.5 Highspeed Across Providers
| Provider | Input (per 1M) | Output (per 1M) |
|---|---|---|
| MiniMax | — | — |
| Vercel AI Gateway | $0.60 | $2.40 |
| Novita AI | $0.60 | $2.40 |
Pricing
| Type | Price (per 1M) |
|---|---|
| Input tokens | $0.60 |
| Output tokens | $2.40 |
Capabilities
About MiniMax M2.5 Highspeed
MiniMax M2.5 Highspeed is MiniMax's inference-optimized variant of M2.5, released simultaneously in February 2026. It delivers identical intelligence and outputs to standard M2.5 through a specialized inference engine at lower latency. The model supports a 204,800-token context window, 131,072-token max output, function calling, structured output, and reasoning. API model ID: MiniMax-M2.5-highspeed. It is designed for latency-sensitive interactive applications and automated agent pipelines.