Microsoft released MAI Image 2.5 Pro and MAI Voice 2 Flash on July 23 24, 2026, cutting GPU costs up to 84% in PowerPoint versus GPT Image 2 and up to 89% in Dynamics 365 Contact Center versus OpenAI voice models, mar...

Create a landscape editorial hero image for this Studio Global article: Search & fact-check with cited sources for What two new in-house AI models did Microsoft launch on Thursday, how are they being deployed acr. Article summary: Here is a comprehensive answer based on the strongest available sources.. Topic tags: general, general web, user generated. Style: premium digital editorial illustration, source-backed research mood, clean composition, high detail, modern web publication hero. Use reference image context only for broad subject, composition, and topical grounding; do not copy the exact image. Avoid: logos, brand marks, copyrighted characters, real person likenesses, fake screenshots, UI text, readable text, watermarks, charts with fake numbers, clickbait thumbnails, icons, and tiny thumbnail layouts. Make it useful as an illustrative visual, not as factual evidence.
On July 23–24, 2026, Microsoft released two new variants of its in-house MAI AI models — MAI-Image-2.5-Pro and MAI-Voice-2-Flash — and deployed them across Bing Image Creator, PowerPoint, OneDrive, and Dynamics 365 Contact Center, cutting GPU costs by up to 89% compared to OpenAI's equivalent models . The launch completes a months-long strategy to replace OpenAI models in Microsoft's highest-volume consumer and enterprise products with models built from scratch on Microsoft's own Maia 200 silicon. Here is what the two models are, where they are being used, how much they save, and what this means for Microsoft's long-term AI infrastructure.
Both models entered public preview through Microsoft Foundry on July 24 .
Microsoft is rolling the new models into several flagship products immediately :
Microsoft states that MAI models are now deployed across more than half of its products, with further replacements underway in Excel and Outlook .
Microsoft provided specific GPU compute cost reduction percentages — meaning the savings on inference infrastructure when running its own models versus licensing or hosting OpenAI's equivalents. These are not API pricing differences but actual hardware-level savings :
These reductions are significant enough to change Microsoft's unit economics for AI features across its Office and cloud suite.
The July 24 releases are specialized variants within a broader seven-model MAI family announced at Microsoft Build in June 2026 :
MAI-Image-2.5-Pro adds a premium tier for maximum image quality, while MAI-Voice-2-Flash adds a speed-optimized tier for high-volume, low-latency voice . The entire family is trained on Microsoft's own Maia 200 custom AI accelerators without distillation from third-party models, with a claimed 1.4x efficiency advantage from the hardware-software co-design
.
Microsoft is executing a deliberate multi-phase strategy to build full-stack AI self-sufficiency :
Studio Global AI
Use this topic as a starting point for a fresh source-backed answer, then compare citations before you share it.
Microsoft released MAI Image 2.5 Pro and MAI Voice 2 Flash on July 23 24, 2026, cutting GPU costs up to 84% in PowerPoint versus GPT Image 2 and up to 89% in Dynamics 365 Contact Center versus OpenAI voice models, mar...