Mistral Medium 3.5 on Vercel AI Gateway

Mistral Medium · MistralAI

ServerlessOpen Weights

Last refreshed 2026-06-29. Next refresh: weekly.

Why use Mistral Medium 3.5 on Vercel AI Gateway?

Vercel AI Gateway offers Mistral Medium 3.5 with pay-as-you-go pricing at $1.50/1M input tokens. Vercel AI Gateway is a unified AI proxy providing a single OpenAI-compatible API endpoint to 275+ models from 25+ providers including Anthropic, OpenAI, Google, Meta, Mistral, DeepSeek, xAI, Alibaba, Amazon, ByteDance, Cohere, MiniMax, MoonshotAI, KwaiPilot, Black Forest Labs, Recraft, Voyage AI, NVIDIA, and more.

Compare Mistral Medium 3.5 across 3 providers to find the best fit for your use case
Input / 1M
$1.50
Output / 1M
$7.50
Cache
Not sourced
Batch
Not sourced

Setup recipe

Python + curl
Install
pip install openai
Auth
export AI_GATEWAY_API_KEY=...
Call
import os
from openai import OpenAI
client = OpenAI(
    api_key=os.environ["AI_GATEWAY_API_KEY"],
Model ID
mistral/mistral-medium-3.5

Request example

import os
from openai import OpenAI

client = OpenAI(
    api_key=os.environ["AI_GATEWAY_API_KEY"],
    base_url="https://ai-gateway.vercel.sh/v1"
)
response = client.chat.completions.create(
    model="mistral/mistral-medium-3.5",
    messages=[{"role": "user", "content": "Hello"}]
)
print(response.choices[0].message.content)

Gotchas

  • Use provider model ID "mistral/mistral-medium-3.5", not the LLMReference slug "mistral-medium-3.5".
  • creator/model-name e.g. kwaipilot/kat-coder-pro-v2
  • The examples expect AI_GATEWAY_API_KEY; rename it only if your application config maps the new variable.

Compare Mistral Medium 3.5 Across Providers

ProviderInput (per 1M)Output (per 1M)
Mistral AI Studio$1.50$7.50
OpenRouter$1.50$7.50
Vercel AI Gateway$1.50$7.50

Pricing

TypePrice (per 1M)
Input tokens$1.50
Output tokens$7.50

Capabilities

VisionMultimodalReasoningJSON / Tool useStructured Outputs

About Mistral Medium 3.5

Mistral's 128B flagship merged model, released April 29, 2026. Open weights available on HuggingFace since ~May 22, 2026 under a Modified MIT License (free for companies with <$20M/month revenue). Dense 128B architecture with 256k context window, configurable reasoning effort, vision encoder for variable image sizes, and native function calling. Achieves 91.4% on τ³-Telecom and 77.6% on SWE-Bench Verified. Strong multilingual support across 24+ languages. Replaces Mistral Medium 3.1, Magistral, and Devstral 2 as Mistral's primary merged model. Available via Mistral API at $1.50/$7.50 per million tokens (input/output) and via NVIDIA NIM/endpoints.

Get Started

Model Specs

Released2026-04-29
Parameters128B
Context262k
ArchitectureDecoder Only