
Google’s cheap model is now two versions ahead of its flagship
THE SO WHAT
Google’s 3.7 Flash undercutting on price at $0.75 per million tokens while leapfrogging the delayed 3.5 Pro suggests the near-term race is for fast, cheap, “good enough” models, not just frontier peaks. If you’re building AI features, design for model swapability and assume your default will be a high-throughput, low-cost tier.
READ THE SOURCE
MORE FROM THE WIRE
Applied AIMicrosoft’s Clippy-like Mico character is no longer the face of Copilot
Microsoft is de-emphasizing mascot UX in favor of more neutral, enterprise-safe interfaces for Copilot. If you’re shipping AI assistants into professional workflows, lean into clarity and trust cues over playful avatars — brand tone is now a governance decision.
Applied AIOpenAI’s Revenue Run Rate Tops $40 Billion Ahead of IPO
A $40B run rate roughly doubling in a year cements LLM infra and assistants as a core spend category, not an experiment line item. Expect procurement, pricing, and dependency risk around this vendor to become board-level topics as an IPO forces more disclosure and scrutiny.
Applied AIMark Zuckerberg’s AI Manifesto Is 6,500-Words—and Barely Says Anything
Long-form AI manifestos from platform CEOs are now part of the signaling game — but operators should read them as political and regulatory positioning, not product roadmaps. The real information is what ships and where capex goes, not the narrative wrapper.
Applied AIWriter introduces new AI model and upgraded harness to contain token costs
Writer building a post-trained variant on GLM-5.2 with a custom harness is a clear tell — enterprises want deployment-ready models with predictable token economics more than raw benchmark wins. If you’re buying or building, model choice now has to be tied directly to cost-per-workflow, not just capability.