In a lengthy post on X, the Microsoft CEO laid out his views on the dangers posed by highly advanced AI models and how to address those risks. Nadella says we can no longer accept a world in which AI is treated as a “set of nested black boxes” whose advice and actions we simply accept or reject. It calls for building a more transparent system where models can be contained, observed, and leave behind “tamper-proof, human-readable evidence.”
Many of their recommendations align with what we’ve heard from others in the industry: timely incident disclosure, independent audits, verifiable data, and containment. It is on this last point where he seems to go a little further than others in the field, saying:
We must assume that a model is compromised and contain it from the beginning. Think of it as an emergency brake. An authorized person should always be able to pause or turn off a model mid-task. More advanced models will require more advanced containment technologies that we must standardize.
Unfortunately, Nadella also refers to AI as “superintelligence” throughout the post.