He says to assume a model is compromised
TechCrunch reports that Microsoft CEO Satya Nadella posted on X on Saturday morning that it is time to step back and assess the trust architecture of AI.
He wrote that we cannot treat super intelligence as a set of nested black boxes and simply accept or reject its recommendations, answers, and actions. TechCrunch notes that super intelligence is the Trump administration's preferred term. This article does not open the X post. The quotations below come from TechCrunch.
[1]
The brake means someone can stop a task
Nadella describes the approach as separating the model from the harness that orchestrates its work, and externalizing controls and safeguards. He also called for every meaningful model action to be documented with tamper-proof human readable evidence. An authorized person should always be able to pause or shut down a model mid-task.
His line is: we must assume a model is compromised and contain it from the start. Think of it like an emergency brake.
[1]What this report does not include
TechCrunch does not quote implementation steps, a product, or a timetable. It places the post against a background: leading AI companies have acknowledged incidents where they seemed to lose control of models, and Anthropic CEO Dario Amodei published a plan for more cautious development. The details of those incidents are not in this article.
[1]