Last refreshed 2026-05-16. Next refresh: weekly.
Why use Mixtral 8x7B on Microsoft Foundry?
Microsoft Foundry offers Mixtral 8x7B with pay-as-you-go pricing at $0.27/1M input tokens. Microsoft Foundry is a unified Azure platform-as-a-service offering for enterprise AI operations, model builders, and application development.
Compare Mixtral 8x7B across 18 providers to find the best fit for your use caseSetup recipe
Docs fallbackUse the provider REST API or SDKCreate a provider API keymodel: mixtral-8x7bmixtral-8x7bRequest example
mixtral-8x7b.Gotchas
No curated gotchas have been sourced for this exact provider/model route yet.
Compare Mixtral 8x7B Across Providers
| Provider | Input (per 1M) | Output (per 1M) |
|---|---|---|
| Databricks Foundation Model Serving | $0.50 | $1.00 |
| NVIDIA NIM | — | — |
| GCP Vertex AI | $0.40 | $1.20 |
| AWS Bedrock | $0.45 | $0.70 |
| OctoAI API (Deprecated) | — | — |
Pricing
| Type | Price (per 1M) |
|---|---|
| Input tokens | $0.27 |
| Output tokens | $0.27 |
Capabilities
No model capability flags are currently sourced.
About Mixtral 8x7B
Mixtral 8x7B, developed by Mistral AI, features a cutting-edge Mixture of Experts (MoE) architecture, utilizing eight experts with seven billion parameters each, yielding a total of 46.7 billion parameters. This architecture activates only two experts per token, allowing for efficient processing and a 6x faster inference rate compared to Llama 2 70B. The model excels in performance, surpassing Llama 2 70B and competing with GPT-3.5 on numerous benchmarks. It supports multiple languages and can handle context up to 32,000 tokens, enhancing understanding of lengthy text.