Microsoft CEO Satya Nadella calls for stronger safety controls for advanced AI systems, saying companies should “step back and assess the trust architecture” behind their models. In comments shared on Oct. 10, 2026, he argues that AI models should be separated from the systems that orchestrate their actions, with safeguards placed outside the model itself.

Nadella says organizations should initially assume an AI model could be compromised, limit its capabilities accordingly, and ensure an authorised human can stop the system mid-task. He also calls for tamper-resistant, human-readable records of significant actions, alongside tools that can trace agent behavior to specific parts of code.

Different outlets frame the same theme as “emergency brake” controls, trust and accountability architecture, and the need for timely disclosure when systems fail. Several reports also place his remarks alongside broader industry discussions following incidents in which companies acknowledged they had lost control of AI models.