Mistral Large 4
- Family
- Mistral Large
- Released
- 2026-10-06
- Context
- 1.05m
- Parameters
- 1.05T (49B active)
- Architecture
- Mixture of Experts
- Specialization
- general
- Openness
- Proprietary
- License
- ProprietaryCommercial use: conditional
- Weights
- Not released
- Code
- Unknown
- Training
- Multi-stage
Cheapest of 3 routes · Mistral AI Studio · cache read $0.070
About
Mistral Large 4 (API id mistral-large-4-0, public preview launched 6 October 2026) is Mistral AI's flagship open-weight multimodal Mixture-of-Experts model: 1.05T total parameters, 49B active, and a 1.6B vision encoder, with a 1M-token context window on first-party docs. It accepts text and image input and returns text. First-party docs list structured outputs, function calling, document QnA, prefix, chat completions, batching, agents/conversations, and built-in tools. Standard API pricing on docs.mistral.ai/inference/pricing is $0.68 / $0.07 cached input / $2.09 output per 1M tokens.
Provider price ladder
Compare all 3Compare API pricing across 3 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Cache | Route |
|---|---|---|---|---|
| Mistral AI Studio | $0.680 | $2.09 | read $0.070 | Serverless |
| OpenRouter | $0.680 | $2.09 | read $0.070 | Serverless |
| Vercel AI Gateway | $0.680 | $2.09 | read $0.070 | Serverless |
Available via routers & gateways(10)
LiteLLM
GatewayOpen-source Python SDK and proxy server that unifies 100+ LLM APIs behind a single OpenAI-compatible interface, with load balancing, cost tracking, and configurable failover.
OpenRouter
HybridUnified hybrid gateway to 400+ models from 60+ providers via a single OpenAI-compatible API, with optional auto-routing that selects the best model per prompt.
Portkey
GatewayProduction AI gateway routing to 1,600+ LLMs with failover, load balancing, semantic caching, and guardrails; Apache 2.0 core is fully self-hostable with the complete feature set.
AIRouter
RouterCommercial LLM router that analyzes incoming requests and routes to the optimal model for cost/quality/latency via a drop-in OpenAI-compatible API, with a privacy-preserving embedding mode that avoids sending prompt content.
Martian
RouterAI-powered LLM router that analyzes each prompt in real-time to select the optimal model, targeting 20–97% cost reduction while maintaining quality; San Francisco startup reportedly nearing $1.3B valuation.
Neutrino AI
RouterCommercial LLM router that dynamically routes each query to the best-suited model with load balancing and fallback handling, charging 3% of underlying AI spend.
Capabilities
Benchmark peer barsfor Coding
No task-mapped benchmark peers are available for this model yet.
Migration checks
No linked migration route is available for this model yet.
API versions
mistral-large-4-0mistral-large-4ML4mistralai/mistral-large-4-0mistral/mistral-large-4Rankings & picks(2)
Cheapest of 3 routes · Mistral AI Studio · cache read $0.070