
California is deciding who may verify AI, and one investigation already cost $400,000 in tokens
THE SO WHAT
Verification is becoming a licensed profession with a real cost structure—California’s SB 813 and METR’s ~$400,000 token burn show evals are neither cheap nor purely academic. If you build or buy frontier models, budget for third‑party verification as a line item and start mapping which certifiers will actually matter to your regulators and customers.
READ THE SOURCE
MORE FROM THE WIRE
Applied AIAstra appears to think without showing its work, and the people arguing about it co-wrote the warning
If frontier models are doing more of their reasoning off-text, your eval and oversight stack is already out of date. Treat monitorability as a first-class requirement in vendor selection and internal research — not a nice-to-have interpretability feature.
Applied AIOpenAI confirms ‘wiki incident,’ says it’s ‘working on a framework’ for more disclosure
Agentic deployments are now creating real-world governance incidents, and disclosure frameworks are becoming part of the product surface. If you're piloting autonomous agents, assume post-incident transparency expectations will rise and design logging, kill-switches, and comms protocols now.
OpenAI says it will change how it informs the public when its AI agents go off the rails
Incident disclosure around agent misbehavior is becoming part of the product surface, not just a compliance afterthought. If you’re deploying agentic systems, assume customers and regulators will expect a playbook for when things go wrong that looks closer to breach disclosure than bug reporting.
Applied AIMinnesota nudification law survives xAIs request for a legal injunction
A $500,000-per-incident exposure for AI-enabled ‘nudification’ sets a concrete price tag on misuse risk. Any team shipping image models or tooling now needs jurisdiction-level abuse mapping and logging, not just generic terms of service.