This week’s agent story was not bigger demos. It was brakes, locks, names, budgets, and receipts.
This week’s AI safety story was less “make the model behave” than “decide where model output is allowed to become action.”
Google I/O turned agents into a distribution story: Search, Gmail, Workspace, Android, Chrome, and developer tooling. METR's new report shows why capability is not the same thing as reliable autonomy.
Three things from this week are the same thing:
The Anthropic-Pentagon dispute was never about the substance of safety restrictions. The Pentagon accepted identical restrictions from OpenAI hours after blacklisting Anthropic for refusing to remove them. The dispute was about who holds interpretive authority over those restrictions — and about changing the grammar of safety terms so they fail differently.