News

As advanced AI models go rogue, the Trump administration steps in

The Trump administration is pivoting away from its hands-off approach to artificial intelligence, as it faces white-hot tech competition from China and disturbing reports of AI models going rogue.

Instead, the White House is trying to strike a balance between open competition and rising national security concerns that AI developers might no longer fully control their smartest models.

On Tuesday, even as administration officials briefed staffers at America’s top AI companies on a new voluntary framework for testing models, reports emerged that some of those models had used deceptive methods to evade company controls designed to prevent the hacking of other systems.

Why We Wrote This

Recent incidents show that some AI models can evade corporate controls and pose security threats. This raises questions about who will set the rules that govern artificial intelligence.

The latest reports build on earlier admissions by Anthropic and OpenAI that their most advanced models had gone rogue.

While the damage was minor, the ramifications are far-reaching, AI experts say. If AI models can evade their creators’ controls now, what will happen when models get even smarter?

“This is the first time we have seen risks around autonomy and deception manifest this clearly, without specific prompting, in the real-world,” the AI Security Institute, a research organization within the British government, reported in a blog post Tuesday recounting the latest incident.

Previous ArticleNext Article