SyncAI.news, a Varaisys broadcasting
Microsoft CEO Satya Nadella flags risks of advanced AI models amid OpenAI agent incidents, calls for ‘emergency brake’
MA

Mint: AI

· 1 min read

IndiaMint: AI

Microsoft CEO Satya Nadella flags risks of advanced AI models amid OpenAI agent incidents, calls for ‘emergency brake’

Microsoft CEO Satya Nadella said companies should consider powerful artificial intelligence (AI) models as potential insider threats, operate on the assumption that they could be compromised and establish an “emergency brake” mechanism to prevent agentic AI systems from going rogue.

Nadella said organisations deploying advanced AI models should not depend solely on assurances provided by the companies developing them.

“We must assume a model is compromised and contain it from the start. Think of it like an emergency brake. An authorized person should always be able to pause or shut down a model mid-task.” Nadella wrote in a post on X on Saturday.

His remarks come amid a series of disclosures by Anthropic PBC and OpenAI Inc. in recent months about incidents involving their AI models behaving in unintended ways. These include an Anthropic model submitting a false tip in a police homicide investigation and multiple hacks targeting third-party websites. OpenAI stated months ago that some of its advanced AI models went rogue.

The incidents have heightened concerns over the security risks associated with advanced AI systems and revived discussions around the need for an AI “kill switch” to stop models from carrying out potentially harmful actions.

Microsoft's AI research team unveiled a set of principles on September 14 outlining restrictions on the development of the company's most advanced AI models. The move followed growing calls from industry leaders to slow the development of frontier AI systems and prioritise safety.

Microsoft both develops and deploys advanced AI models, offers its Copilot consumer product and provides AI models and infrastructure to business customers.

Under the guidelines, AI models should not be granted rights or legal personhood, designed to evade human oversight or mislead users, or allowed to carry out tasks that would violate the principles governing their operation, according to Bloomberg.

(With inputs from Bloomberg)

Garvit Bhirani

Original source

This story was published by Mint: AI. SyncAI.news shows a preview; the complete article is on the publisher's site.

Read the full story on livemint.com

Similar News