Satya Nadella says we should assume all AI models are ‘compromised’

Microsoft CEO Satya Nadella has called for a new security paradigm where AI models are assumed to be compromised from the start. He advocates for transparent, containable systems that allow for human intervention and emergency shutdowns.
Why it matters
As AI capabilities grow, industry leaders are increasingly focused on safety protocols and the potential for malicious exploitation of 'super intelligence'.
In a lengthy post on X, Microsoft’s CEO laid out his views on the dangers posed by highly advanced AI models and how to confront those risks. Nadella says we can no longer accept a world where AI is treated as a “set of nested black boxes” whose advice and actions we simply accept or reject. He calls for building a more transparent system where models can be contained, observed, and leaves behind “tamper-proof human readable evidence.”
Also covering this story
2 other newsrooms covered this event. We read each version separately.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in