LLM Reference
Vultr

Mixtral 8x7B on Vultr

Mixtral · MistralAI

ServerlessOpen Source

Last refreshed 2026-06-29. Next refresh: weekly.

Why use Mixtral 8x7B on Vultr?

Vultr offers Mixtral 8x7B with pay-as-you-go pricing at $0.55/1M input tokens. Vultr is a cloud infrastructure company headquartered in West Palm Beach, Florida.

Compare Mixtral 8x7B across 18 providers to find the best fit for your use case
Input / 1M
$0.55
Output / 1M
$2.75
Cache
Not sourced
Batch
Not sourced

Setup recipe

Docs fallback
Install
Use the provider REST API or SDK
Auth
Create a provider API key
Call
model: mixtral-8x7b
Model ID
mixtral-8x7b

Request example

Curated snippets for this provider are not sourced yet. Use Vultr documentation with model ID mixtral-8x7b.

Gotchas

No curated gotchas have been sourced for this exact provider/model route yet.

Compare Mixtral 8x7B Across Providers

ProviderInput (per 1M)Output (per 1M)
Databricks Foundation Model Serving$0.50$1.00
NVIDIA NIM——
GCP Vertex AI$0.40$1.20
AWS Bedrock$0.45$0.70
OctoAI API (Deprecated)——
View all 18 providers →

Pricing

TypePrice (per 1M)
Input tokens$0.55
Output tokens$2.75

Capabilities

No model capability flags are currently sourced.

About Mixtral 8x7B

Mixtral 8x7B, developed by Mistral AI, features a cutting-edge Mixture of Experts (MoE) architecture, utilizing eight experts with seven billion parameters each, yielding a total of 46.7 billion parameters. This architecture activates only two experts per token, allowing for efficient processing and a 6x faster inference rate compared to Llama 2 70B. The model excels in performance, surpassing Llama 2 70B and competing with GPT-3.5 on numerous benchmarks. It supports multiple languages and can handle context up to 32,000 tokens, enhancing understanding of lengthy text.

Get Started

Model Specs

Released2023-12-11
Parameters8x7B
Context32k
ArchitectureMixture of Experts
Knowledge cutoff2023-12