Both models entered public preview through Microsoft Foundry on July 24 .
Microsoft is rolling the new models into several flagship products immediately :
Microsoft states that MAI models are now deployed across more than half of its products, with further replacements underway in Excel and Outlook .
Microsoft provided specific GPU compute cost reduction percentages — meaning the savings on inference infrastructure when running its own models versus licensing or hosting OpenAI's equivalents. These are not API pricing differences but actual hardware-level savings :
| Model | Use Case | GPU Cost Reduction vs. OpenAI | Source |
|---|---|---|---|
| MAI-Image-2.5-Pro | PowerPoint (vs. GPT-Image-2) | Up to 84% | |
| MAI-Voice-2-Flash | Dynamics 365 Contact Center | Up to 89% | |
| MAI-Voice-2-Flash | General (vs. MAI-Voice-2) | 32% more cost-effective (plus 2x speed) |
These reductions are significant enough to change Microsoft's unit economics for AI features across its Office and cloud suite.
The July 24 releases are specialized variants within a broader seven-model MAI family announced at Microsoft Build in June 2026 :
MAI-Image-2.5-Pro adds a premium tier for maximum image quality, while MAI-Voice-2-Flash adds a speed-optimized tier for high-volume, low-latency voice . The entire family is trained on Microsoft's own Maia 200 custom AI accelerators without distillation from third-party models, with a claimed 1.4x efficiency advantage from the hardware-software co-design .
Microsoft is executing a deliberate multi-phase strategy to build full-stack AI self-sufficiency :