Last refreshed 2026-10-06. Next refresh: weekly.
Why use MAI-Voice-2.1-Flash on Microsoft Foundry?
Microsoft Foundry offers MAI-Voice-2.1-Flash with competitive pricing. Microsoft Foundry is a unified Azure platform-as-a-service offering for enterprise AI operations, model builders, and application development.
Compare MAI-Voice-2.1-Flash across 2 providers to find the best fit for your use caseInput / 1M
-
Output / 1M
-
Cache
Not sourced
Batch
Not sourced
Setup recipe
Docs fallbackInstall
Use the provider REST API or SDKAuth
Create a provider API keyCall
model: MAI-Voice-2.1-FlashModel ID
MAI-Voice-2.1-FlashRequest example
Curated snippets for this provider are not sourced yet. Use Microsoft Foundry documentation with model ID
MAI-Voice-2.1-Flash.Gotchas
- Use provider model ID "MAI-Voice-2.1-Flash", not the LLMReference slug "mai-voice-2-1-flash".
Compare MAI-Voice-2.1-Flash Across Providers
| Provider | Input (per 1M) | Output (per 1M) |
|---|---|---|
| Microsoft Foundry | — | — |
| Vercel AI Gateway | — | — |
Capabilities
Audio
About MAI-Voice-2.1-Flash
MAI-Voice-2.1-Flash is Microsoft AI's low-latency text-to-speech model for high-volume voice agents, launched 1 October 2026 alongside MAI-Voice-2.1. It covers the same 23 languages and cross-language voices with about 150 ms end-to-end latency and supports voice cloning. Priced at $15 per 1M characters.
Get Started
Model Specs
Released2026-10-01
ArchitectureAudio / Speech