Last refreshed 2026-10-06. Next refresh: weekly.
Why use MAI-Voice-2.1-Flash on Vercel AI Gateway?
Vercel AI Gateway offers MAI-Voice-2.1-Flash with competitive pricing. Vercel AI Gateway is a unified AI proxy providing a single OpenAI-compatible API endpoint to 275+ models from 25+ providers including Anthropic, OpenAI, Google, Meta, Mistral, DeepSeek, xAI, Alibaba, Amazon, ByteDance, Cohere, MiniMax, MoonshotAI, KwaiPilot, Black Forest Labs, Recraft, Voyage AI, NVIDIA, and more.
Compare MAI-Voice-2.1-Flash across 2 providers to find the best fit for your use caseInput / 1M
-
Output / 1M
-
Cache
Not sourced
Batch
Not sourced
Setup recipe
Python + curlInstall
pip install openaiAuth
export AI_GATEWAY_API_KEY=...Call
import os
from openai import OpenAI
client = OpenAI(
api_key=os.environ["AI_GATEWAY_API_KEY"],Model ID
microsoft/mai-voice-2.1-flashRequest example
import os
from openai import OpenAI
client = OpenAI(
api_key=os.environ["AI_GATEWAY_API_KEY"],
base_url="https://ai-gateway.vercel.sh/v1"
)
response = client.chat.completions.create(
model="microsoft/mai-voice-2.1-flash",
messages=[{"role": "user", "content": "Hello"}]
)
print(response.choices[0].message.content)Gotchas
- Use provider model ID "microsoft/mai-voice-2.1-flash", not the LLMReference slug "mai-voice-2-1-flash".
- creator/model-name e.g. kwaipilot/kat-coder-pro-v2
- The examples expect AI_GATEWAY_API_KEY; rename it only if your application config maps the new variable.
Compare MAI-Voice-2.1-Flash Across Providers
| Provider | Input (per 1M) | Output (per 1M) |
|---|---|---|
| Microsoft Foundry | — | — |
| Vercel AI Gateway | — | — |
Capabilities
Audio
About MAI-Voice-2.1-Flash
MAI-Voice-2.1-Flash is Microsoft AI's low-latency text-to-speech model for high-volume voice agents, launched 1 October 2026 alongside MAI-Voice-2.1. It covers the same 23 languages and cross-language voices with about 150 ms end-to-end latency and supports voice cloning. Priced at $15 per 1M characters.
Get Started
Model Specs
Released2026-10-01
ArchitectureAudio / Speech