Skip to main content
Microsoft Foundry
Microsoft
MicrosoftMicrosoft's AI Superintelligence team builds enterprise-class AI models spanning reasoning, coding, image, voice, and transcription.

Overview

The latest MAI models are trained from scratch with zero distillation on enterprise-grade, commercially licensed data, providing a clean lineage that is deployable into production with confidence. The flagship MAI-Thinking-1, a 35B-parameter reasoning model with a 256K context window, showcases the ability to compete with frontier models on tough benchmarks like SWE-Bench Pro at a fraction of the size and cost. Alongside it sit fast, inference-efficient models for coding and a chart-topping image model that already ranks among the best for editing.

Why Microsoft Models on Foundry

MAI models are coming to Microsoft Foundry, where they run with first-party managed compute, granular quota, and zero data egress. The real differentiator is Microsoft Frontier Tuning: using reinforcement-learning environments, you adapt MAI models to your own data and workflows within your compliance boundary, so the resulting model stays yours.

Models (91)
Microsoft AI
MAI-Voice-2text-to-speech
Input typetext
Output typeaudio
Preview
Microsoft AI
MAI-Transcribe-1.5automatic-speech-recognition
Input typeaudio
Output typetext
Preview
Microsoft AI
MAI-Image-2.5text-to-image
Input typetext, image
Output typeimage
Context window131.072k
Token limits4.096k output
Preview
Microsoft AI
MAI-Image-2.5-Flashtext-to-image
Input typetext, image
Output typeimage
Context window131.072k
Token limits4.096k output
Preview
Microsoft AI
MAI-Thinking-1chat-completion
Input typetext
Output typetext
Context window256k
Token limits64k output
Preview
Microsoft AI
MAI-Image-2etext-to-image
Input typetext
Output typeimage
Context window131.072k
Token limits4.096k output
Deprecated
Microsoft AI
MAI-Voice-1text-to-speech
Input typetext
Output typeaudio
Preview
Microsoft AI
MAI-Transcribe-1automatic-speech-recognition
Input typeaudio
Output typetext
Preview
model-router
model-routerchat-completion
Input typetext, image
Output typetext
Context window1048.576k
Token limits32.768k output
Generally available (GA)
EvoDiff
EvoDiffprotein-sequence-generation
Input typetext
Output typetext
Token limits4.096k output
Preview
Phi-4-reasoning
Phi-4-reasoningchat-completion
Input typetext
Output typetext
Context window32.768k
Token limits32.768k output
Preview
Phi-4-mini-reasoning
Phi-4-mini-reasoningchat-completion
Input typetext
Output typetext
Context window128k
Token limits128k output
Preview
Phi-4-mini-instruct
Phi-4-mini-instructchat-completion
Input typetext
Output typetext
Context window128k
Token limits4.096k output
Preview
Phi-4-multimodal-instruct
Phi-4-multimodal-instructchat-completion
Input typeaudio, image, text
Output typetext
Context window128k
Token limits4.096k output
Preview
Phi-4
Phi-4chat-completion
Input typetext
Output typetext
Context window16.384k
Token limits16.384k output
Preview
financial-reports-analysis-v2
financial-reports-analysis-v2chat-completion
Input typetext
Output typetext
Context window16.384k
Token limits16.384k output
Preview
supply-chain-trade-regulations-v2
supply-chain-trade-regulations-v2chat-completion
Input typetext
Output typetext
Context window16.384k
Token limits16.384k output
Preview
Muse
Museimage-to-image
Input typeimage
Output typeimage
Generally available (GA)
Prov-GigaPath
Prov-GigaPathimage-feature-extraction
Preview
Azure-Language-Text-Analytics-for-Health
Azure-Language-Text-Analytics-for-Healthhealth-entity-extraction
Input typetext
Output typetext
Generally available (GA)
Azure-Speech-Text-to-speech
Azure-Speech-Text-to-speechtext-to-speech
Input typetext
Output typeaudio
Generally available (GA)
Phi-3-vision-128k-instruct
Phi-3-vision-128k-instructchat-completion
Generally available (GA)
GigaTIME
GigaTIMEimage-to-image
Input typeimage
Output typeimage
Generally available (GA)
Aurora-1.5
Aurora-1.5environmental-forecasting
Input typedata
Output typedata
Preview
Azure-Language-Document-PII-redaction
Azure-Language-Document-PII-redactiondocument-pii-extraction
Input typetext
Output typetext
Generally available (GA)
qwen2.5-0.5b-instruct-generic-cpu
qwen2.5-0.5b-instruct-generic-cpuchat-completion
Input typetext
Output typetext
Token limits2.048k output
Generally available (GA)
fmv-grounding-detection
fmv-grounding-detectionimage-analysis
Input typetext, image
Output typetext
Preview
microsoft-Orca-2-7b
microsoft-Orca-2-7btext-generation
Generally available (GA)
Aurora
Auroraenvironmental-forecasting
Input typedata
Output typedata
Preview
Fara1.5-27B
Fara1.5-27Bweb-agent-tasks
Input typeimage, text
Output typetext
Generally available (GA)
MedImageInsight
MedImageInsightembeddings
Input typeimage, text
Output typeembeddings
Preview
qwen2.5-0.5b-instruct-generic-gpu
qwen2.5-0.5b-instruct-generic-gpuchat-completion
Input typetext
Output typetext
Token limits2.048k output
Generally available (GA)
Phi-4-reasoning-plus-onnx
Phi-4-reasoning-plus-onnxchat-completion
Input typetext
Output typetext
Preview
MagenticBrain-14B
MagenticBrain-14Bchat-completion
Input typetext
Output typetext
Context window32.768k
Token limits4.096k output
Preview
Azure-Speech-Speech-Translation
Azure-Speech-Speech-Translationtranslation
Input typeaudio
Output typetext, audio
Generally available (GA)
CxrReportGen
CxrReportGenimage-text-to-text
Input typeimage, text
Output typetext
Preview
Microsoft AI
MAI-Code-1.1-Flashchat-completion
Input typetext, image
Output typetext
Context window256k
Token limits64k output
Preview
Azure-Language-Language-detection
Azure-Language-Language-detectiondetect-language
Input typetext
Output typetext
Generally available (GA)
DeepSeek-R1-Distilled-NPU-Optimized
DeepSeek-R1-Distilled-NPU-Optimizedchat-completion
Input typetext
Output typetext
Generally available (GA)
Phi-3-small-8k-instruct
Phi-3-small-8k-instructchat-completion
Input typetext
Output typetext
Context window131.072k
Token limits4.096k output
Generally available (GA)
CxrReportGen-Premium
CxrReportGen-Premiumimage-text-to-text
Input typeimage, text
Output typetext
Preview
Microsoft AI
MAI-Transcribe-2automatic-speech-recognition
Input typeaudio
Output typetext
Preview
microsoft-llava-med-v1.5-mistral-7b
microsoft-llava-med-v1.5-mistral-7bimage-text-to-text
Preview
Boltz-1
Boltz-1embeddings
Input typetext
Output typetext
Preview
Phi-4-mini-reasoning-onnx
Phi-4-mini-reasoning-onnxchat-completion
Input typetext
Output typetext
Preview
Azure-Speech-Speech-to-text
Azure-Speech-Speech-to-textautomatic-speech-recognition
Input typeaudio
Output typetext
Generally available (GA)
Azure-Speech-Voice-Live
Azure-Speech-Voice-Liveconversational-ai
Input typetext, audio
Output typetext, audio
Generally available (GA)
eo-os-object-detection
eo-os-object-detectionimage-analysis
Input typeimage
Output typetext
Preview
Microsoft AI
MAI-Image-2.5-Protext-to-image
Input typetext, image
Output typeimage
Context window131.072k
Token limits4.096k output
Preview
MatterGen
MatterGenmaterials-design
Input typetext
Output typetext
Preview
microsoft-swinv2-base-patch4-window12-192-22k
microsoft-swinv2-base-patch4-window12-192-22kimage-classification
Generally available (GA)
1