LLM Reference

MAI Models by Microsoft AI

Microsoft AIProprietary
14 models2025–2026Up to 256k ctxFrom $0.36/1M input

Last refreshed 2026-09-04. Next refresh: weekly.

Details

ResearcherMicrosoft AI
LicenseProprietary
Commercial useCommercial use: conditional
Models14
Released2025–2026
Max context256k

Capabilities

Vision5 of 14 models
Multimodal9 of 14 models
Reasoning3 of 14 models
JSON / Tool use2 of 14 models

Links

Website

About

Microsoft AI (MAI) is Microsoft's proprietary model family for Copilot and Azure AI Foundry. The lineup now spans reasoning, coding, image generation/editing, speech synthesis, and transcription models, including MAI-Thinking-1, MAI-Code-1, MAI-Code-1-Flash, MAI-Image-2.5, MAI-Image-2.5-Flash, MAI-Voice-2, and MAI-Transcribe-1.5 alongside the earlier MAI image and speech releases.

Decision facts

Best fit
image generationspeech recognitionaudio
Capability starting point
MAI-Thinking-1 with 256k context and reasoning and JSON / Tool use
Lowest tracked input
MAI-Transcribe-1 · $0.36/1M · Microsoft Foundry
Closest related family
Claude 3

Current Variants

Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.

14 in view

Use when the workload needs image generation, 32k context, and multimodal inputs.

2026-09image generation32k contextmultimodal inputs

Use when the workload needs speech recognition, multimodal inputs, and audio.

2026-09speech recognitionmultimodal inputsaudio

Use when the workload needs reasoning, 256k context, and JSON / Tool use.

2026-06reasoning256k contextJSON / Tool use
MAI-Code-1Current

Use when the workload needs code.

2026-06code

Use when the workload needs code, 256k context, and reasoning.

2026-06code256k contextreasoning

Use when the workload needs image generation, 32k context, and multimodal inputs.

2026-06image generation32k contextmultimodal inputs

Use when the workload needs image generation, 32k context, and multimodal inputs.

2026-06image generation32k contextmultimodal inputs

Use when the workload needs text to speech and audio.

2026-06text to speechaudio

Use when the workload needs speech recognition, multimodal inputs, and audio.

2026-06speech recognitionmultimodal inputsaudio

Use when the workload needs image generation, 33k context, and multimodal inputs.

2026-04image generation33k contextmultimodal inputs

Use when the workload needs speech recognition, multimodal inputs, and audio.

2026-04speech recognitionmultimodal inputsaudio

Use when the workload needs text to speech, multimodal inputs, and audio.

2026-04text to speechmultimodal inputsaudio

Use when the workload needs image generation and multimodal inputs.

2026-03image generationmultimodal inputs
MAI-DS-R1Current

Use when the workload needs reasoning, 164k context, and 671B parameters.

2025-04reasoning164k context671B parameters

Release Timeline

5 release groups
2026-09
2 current
MAI-Image-2.6-Flash
image generation32k contextmultimodal inputs
Current
MAI-Transcribe-2
speech recognitionmultimodal inputsaudio
Current
2026-06
7 current
Current
MAI-Code-1-Flash
code256k contextreasoning
Current
MAI-Image-2.5
image generation32k contextmultimodal inputs
Current
MAI-Image-2.5-Flash
image generation32k contextmultimodal inputs
Current
MAI-Thinking-1
reasoning256k contextJSON / Tool use
Current
MAI-Transcribe-1.5
speech recognitionmultimodal inputsaudio
Current
MAI-Voice-2
text to speechaudio
Current
2026-04
3 current
MAI-Image-2e
image generation33k contextmultimodal inputs
Current
MAI-Transcribe-1
speech recognitionmultimodal inputsaudio
Current
MAI-Voice-1
text to speechmultimodal inputsaudio
Current
2026-03
1 current
MAI-Image-2
image generationmultimodal inputs
Current
2025-04
1 current
MAI-DS-R1
reasoning164k context671B parameters
Current

Specifications(14 models)

MAI model specifications comparison
ModelReleasedContextParametersVisionMultimodalReasoningJSON / Tool use
MAI-Image-2.6-Flash2026-0932kYesYesNoNo
MAI-Transcribe-22026-09NoYesNoNo
MAI-Thinking-12026-06256k1T total / 35B activeNoNoYesYes
MAI-Code-12026-06NoNoNoNo
MAI-Code-1-Flash2026-06256kNoNoYesYes
MAI-Image-2.52026-0632kYesYesNoNo
MAI-Image-2.5-Flash2026-0632kYesYesNoNo
MAI-Voice-22026-06NoNoNoNo
MAI-Transcribe-1.52026-06NoYesNoNo
MAI-Image-2e2026-0433kYesYesNoNo
MAI-Transcribe-12026-04NoYesNoNo
MAI-Voice-12026-04NoYesNoNo
MAI-Image-22026-03YesYesNoNo
MAI-DS-R12025-04164k671BNoNoYesNo

Available From(1 provider)

Pricing

MAI model pricing by provider
ModelProviderInput / 1MOutput / 1MType
MAI-Transcribe-1Microsoft Foundry$0.36Serverless
MAI-Code-1-FlashMicrosoft Foundry$0.75$4.5Serverless
MAI-Image-2Microsoft Foundry$5$33Serverless
MAI-Voice-1Microsoft Foundry$22Serverless