Microsoft Cuts AI Costs with Custom Maia 200 Chip
Microsoft's custom AI chip, Maia 200, has been deployed in an Iowa data center and is set to be rolled out near Phoenix, Arizona. The second-generation accelerator was unveiled on January 26, 2026, and claims a significant reduction in operational costs compared to Nvidia hardware.
The Maia 200 packs over 140 billion transistors and delivers more than 10 petaFLOPS at FP4 precision and over 5 petaFLOPS at FP8 within a 750-watt thermal envelope. Microsoft claims a 30% improvement in performance per dollar compared to its existing hardware fleet, along with better performance per watt.
The chip was purpose-built for inference, the phase of AI where a trained model actually responds to queries and generates outputs. It's already running OpenAI's GPT-5.2 models and powering Microsoft's own MAI initiatives.
Analyst estimates put the total cost of ownership reduction at 20% to 30% compared to Nvidia GPUs for inference workloads. The internal unit production cost of each Maia 200 is estimated at roughly 30% to 40% of what high-end Nvidia equivalents cost, with those Nvidia chips priced above $30,000 per unit.