Microsoft launches new in-house AI models it says cut costs up to 89% versus OpenAI
THE SO WHAT
If Microsoft can show production workloads running 50–89% cheaper on MAI-Image-2.5-Pro and MAI-Voice-2-Flash than on OpenAI, the internal-vs-partner model debate just became a hard cost line item. Enterprise AI teams should re-open their TCO spreadsheets this week and model scenarios where foundation vendors are also your direct price competitors.
READ THE SOURCE
MORE FROM THE WIRE
Applied AIAmazon shutters its AGI Lab as part of layoffs in its AGI unit earlier this week; the lab's former head David Luan left in February
Shutting a San Francisco AGI Lab focused on agent usefulness suggests Amazon is consolidating research bets into nearer-term product lines. If you’re selling into or building on their stack, expect more emphasis on deployable assistants and infra economics, less on open-ended AGI exploration.
Applied AIAMD is working with Cerebras to combine AMD Helios and Cerebras wafer-scale chips for an AI inference solution to make workloads faster; CBRS closed up 4.86%
AMD teaming up with Cerebras on a combined Helios + wafer-scale inference rack underscores how heterogeneous AI compute is becoming at the system level, not just the chip level. Infra buyers should start architecting for mixed-vendor, mixed-accelerator clusters rather than assuming a single GPU family will dominate their next buildout.
Applied AIServiceNow’s stock falls as a new AI threat overshadows earnings beat
ServiceNow beating earnings while trading down on fear of OpenAI Presence shows how quickly enterprise software is being repriced on the risk of assistant-native workflows. If you run a SaaS business, you now have to assume that horizontal AI assistants will sit between you and the user — design APIs, pricing, and UX for that mediated reality.
Applied AIChatGPT Health Rolls Out to Everyone While OpenAI Stares Down Major Lawsuits
Rolling out a health-branded assistant while facing lawsuits over allegedly fatal advice moves AI triage from a UX question to a liability and governance question. If you’re deploying AI into any regulated or safety-adjacent domain, you need explicit guardrails, escalation paths, and insurance in place before usage scales, not after.