Using MAI-Transcribe-2-Streaming on Microsoft Foundry
Implementation guide · MAI · Microsoft AI
Serverless
Microsoft Foundry exposes MAI-Transcribe-2-Streaming through model ID MAI-Transcribe-2-Streaming. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.
Last refreshed 2026-10-06. Next refresh: weekly.
Quick Start
- 1
- 2Use the Microsoft Foundry SDK or REST API to call
MAI-Transcribe-2-Streaming— see the documentation for request format.
Code Examples
Pricing on Microsoft Foundry
Capabilities
MultimodalAudio
About MAI-Transcribe-2-Streaming
MAI-Transcribe-2-Streaming is Microsoft AI's first streaming speech-to-text model, launched 1 October 2026. Audio streams over a WebSocket and transcripts return incrementally while the speaker talks, with partial results that are then finalised. It is offered at an introductory $0.54 per hour of audio through the end of 2026.
Model Specs
Released2026-10-01
ArchitectureAudio / Speech