
Satya Nadella says we should assume all AI models are ‘compromised’
In a lengthy article on X, the Microsoft CEO outlined his views on the dangers posed by highly advanced AI models and how to deal with them. Nadella argues that we can no longer accept a world in which AI is treated as a “set of nested black boxes” whose advice and actions we simply accept or reject. It calls for building a more transparent system in which patterns can be contained, observed, and leave behind “tamper-proof, human-readable evidence.”
Many of its recommendations match what we’ve heard from others in the industry: rapid disclosure of incidents, independent audits, verifiable data, and containment. It is on this last point that he seems to go a little further than some others in the field, saying:
We must assume that a model is compromised and contain it from the start. Think of it like an emergency brake. An authorized person must always be able to pause or stop a model mid-job. More advanced models will require more advanced containment technologies that we will need to standardize on.
Unfortunately, Nadella also refers to AI as “super intelligence” throughout his article.
Gn bussni