LLM Reference
MiniMax

Using MiniMax M2.7 Highspeed on MiniMax

Implementation guide · MiniMax M2 · MiniMax

ServerlessOpen Source

MiniMax exposes MiniMax M2.7 Highspeed through model ID MiniMax-M2.7-highspeed. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-06-15. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at MiniMax and generate an API key.
  2. 2
    Use the MiniMax SDK or REST API to call MiniMax-M2.7-highspeed — see the documentation for request format.

Code Examples

See MiniMax documentation for integration details.

Pricing on MiniMax

Capabilities

ReasoningJSON / Tool useStructured Outputs

About MiniMax M2.7 Highspeed

MiniMax M2.7 Highspeed is the inference-optimized variant of MiniMax M2.7, released simultaneously on March 18, 2026. It reaches 100 tokens per second output speed, about 66% faster than standard M2.7, while preserving identical intelligence and outputs through engine optimization rather than weight changes. It supports a 204,800-token context window, 131,072-token max output, function calling, structured output, and reasoning. API model ID: MiniMax-M2.7-highspeed.

Model Specs

Released2026-03-18
Parameters10B active
Context205k
ArchitectureDecoder Only

Provider

MiniMax
MiniMax