Mixtral 8x7B
- Family
- Mixtral
- Released
- 2023-12-11
- Context
- 32k
- Parameters
- 8x7B
- Architecture
- Mixture of Experts
- Knowledge cutoff
- 2023-12
- Specialization
- general
- Openness
- Open source
- License
- Apache 2.0OSI-approvedCommercial use: permitted
- Weights
- Unknown
- Code
- Unknown
- Training
- Fine-tuned
Cheapest of 18 routes · SiliconFlow
About
Mixtral 8x7B, developed by Mistral AI, features a cutting-edge Mixture of Experts (MoE) architecture, utilizing eight experts with seven billion parameters each, yielding a total of 46.7 billion parameters. This architecture activates only two experts per token, allowing for efficient processing and a 6x faster inference rate compared to Llama 2 70B. The model excels in performance, surpassing Llama 2 70B and competing with GPT-3.5 on numerous benchmarks. It supports multiple languages and can handle context up to 32,000 tokens, enhancing understanding of lengthy text.
Top use-case fit: coding, agents, and build tasks
Coding
Q/$ AClassification
Q/$ BProvider price ladder
Compare all 18Compare API pricing across 4 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Route |
|---|---|---|---|
| SiliconFlow | $0.200 | $0.200 | Serverless |
| Microsoft Foundry | $0.270 | $0.270 | Provisioned |
| Lepton AI API | $0.300 | $0.300 | Serverless |
| Mistral AI Studio | $0.150 | $0.450 | Serverless |
Available via routers & gateways(16)
LiteLLM
GatewayOpen-source Python SDK and proxy server that unifies 100+ LLM APIs behind a single OpenAI-compatible interface, with load balancing, cost tracking, and configurable failover.
OpenRouter
HybridUnified hybrid gateway to 400+ models from 60+ providers via a single OpenAI-compatible API, with optional auto-routing that selects the best model per prompt.
Portkey
GatewayProduction AI gateway routing to 1,600+ LLMs with failover, load balancing, semantic caching, and guardrails; Apache 2.0 core is fully self-hostable with the complete feature set.
AIRouter
RouterCommercial LLM router that analyzes incoming requests and routes to the optimal model for cost/quality/latency via a drop-in OpenAI-compatible API, with a privacy-preserving embedding mode that avoids sending prompt content.
Amazon Bedrock Intelligent Prompt Routing
RouterAWS Bedrock's native intelligent prompt router that routes prompts between Anthropic Claude model tiers (Haiku/Sonnet) based on predicted task complexity, with no extra per-routing charge.
Azure AI Foundry Model Router
RouterMicrosoft Azure AI Foundry's native model router that uses a trained ML model to route each prompt in real time to the optimal Azure-hosted model, with Balanced/Cost/Quality mode selection and automatic failover.
Capabilities
No model capability flags are currently sourced.
Benchmark peer barsfor Coding
Benchmark scores(4)
| Benchmark | Score | Version | Evaluation | Source |
|---|---|---|---|---|
| Google-Proof Q&A | 54.8 | diamondObserved 2026-03-06 | — | Source |
| HellaSwag | 90.9 | 10-shotObserved 2026-03-06 | — | Source |
| HumanEval | 80.5 | pass@1Observed 2026-03-06 | — | Source |
| Massive Multitask Language Understanding | 80.2 | 5-shotObserved 2026-03-06 | — | Source |
Migration checks
No linked migration route is available for this model yet.
Rankings & picks(1)
Compare Mixtral 8x7B with other models
Comparison and alternatives
Browse all comparisons →Show all 39 popular comparisonssorted by 7-day search impressions
Cheapest of 18 routes · SiliconFlow