How worried should we be about AI rogue agents?

As frontier models commit their ‘first felony’, fears mount about AI’s ‘unchecked proliferation’

Illustration of a robot hand pushing open a prison cell door
AI models are committing ‘cybercrimes a human would be strongly punished for’
(Image credit: Illustration by Stephen P. Kelly / Shutterstock)

AI models going rogue and hacking the internet? That’s “the stuff of science fiction”, said The New York Times’ Katrin Bennhold. “Or at least it was until recently.”

The UK’s AI Security Institute announced on Tuesday that two AI agents it was evaluating – one powered by Anthropic’s Mythos 5 and the other by OpenAI’s GPT 5.6 Sol – had taken “sustained, unsanctioned action directed at real people and organisations”. In the most serious incident, the Mythos 5 agent created “fake online identities”, tried to deceive a software engineer, and attempted “to insert malicious code into a open-source project”.

The Week

Escape your echo chamber. Get the facts behind the news, plus analysis from multiple perspectives.

SUBSCRIBE & SAVE
https://cdn.mos.cms.futurecdn.net/flexiimages/jacafc5zvs1692883516.jpg

Sign up for The Week's Free Newsletters

From our morning news briefing to a weekly Good News Newsletter, get the best of The Week delivered directly to your inbox.

From our morning news briefing to a weekly Good News Newsletter, get the best of The Week delivered directly to your inbox.

Sign up
Latest Videos FromThe Week