Microsoft Chairman and CEO Satya Nadella on Saturday called for advanced artificial intelligence systems to be built with containment, independent controls, and an "emergency brake" that lets authorized people pause or shut down a model mid-task.
Nadella outlined his position in a post on X, arguing that the industry must rethink how frontier AI systems are designed and governed before the technology outpaces the safeguards meant to contain it.
"We need to surround non-deterministic models with strong, deterministic system design, human controls, and reliable operating procedures, and establish industry standards where existing ones are insufficient," Nadella wrote.
He added that treating frontier closed and open-weight models as insider risks is one way to construct such a system — a framing that positions powerful AI models as potential threats to be managed from within, rather than tools to be deployed without restriction.
Nadella enumerated several principles he described as central to "observability" in AI system design. They include model diversity, a human-readable record of a model's actions, continuous system testing, independent controls and auditability, containment mechanisms, and mandatory incident disclosure.
His framing of trustworthiness was notably counterintuitive. "The most trustworthy Super Intelligence system will not be the one with the model we trust most," Nadella wrote. "It will be the one that enables us to trust the model the least."
Nadella's remarks arrive amid a period of heightened public debate over AI safety among the industry's most prominent figures. Microsoft co-founder Bill Gates, Anthropic CEO Dario Amodei, OpenAI CEO Sam Altman, and SpaceX CEO Elon Musk have all issued warnings about what they describe as insufficient AI safety protocols and concerns that development is moving faster than the field can responsibly manage.
Those concerns sharpened in recent weeks. An AI researcher resigned from Anthropic last month and accused both Anthropic and OpenAI of "gambling with our lives." Later the same day, an alignment lead at Anthropic focused on AI safety stated there is a greater than 10% chance the technology could "kill all humans" within the next decade.
The backdrop in Washington adds another layer of tension. President Donald Trump has repeatedly dismissed AI extinction risks, emphasizing instead that the United States must maintain its competitive edge over China. Trump recently introduced a new "AI Force," led by Director of National Intelligence Jay Clayton, aimed at facilitating the industry and identifying bad actors — a safety-oriented initiative that nonetheless stops well short of endorsing any slowdown in frontier development.
That divide — between executives urging structural safeguards and an administration prioritizing speed and geopolitical advantage — is unlikely to resolve quickly. Nadella's call for enforceable industry standards and human override mechanisms signals that at least some of the industry's most powerful voices are willing to advocate for internal constraints even as regulatory clarity remains elusive.