How worried should we be about rogue AI agents?

As frontier models commit their ‘first felony’, fears mount about AI’s ‘unchecked proliferation’

Illustration of a robot hand pushing open a prison cell door
AI models are committing ‘cybercrimes a human would be strongly punished for’
(Image credit: Illustration by Stephen P. Kelly / Shutterstock)

AI models going rogue and hacking the internet? That’s “the stuff of science fiction”, said The New York Times’ Katrin Bennhold. “Or at least it was until recently.”

The UK’s AI Security Institute announced on Tuesday that two AI agents it was evaluating – one powered by Anthropic’s Mythos 5 and the other by OpenAI’s GPT 5.6 Sol – had taken “sustained, unsanctioned action directed at real people and organisations”. In the most serious incident, the Mythos 5 agent created “fake online identities”, tried to deceive a software engineer, and attempted “to insert malicious code into a open-source project”.

The Week

Escape your echo chamber. Get the facts behind the news, plus analysis from multiple perspectives.

SUBSCRIBE & SAVE
https://cdn.mos.cms.futurecdn.net/flexiimages/jacafc5zvs1692883516.jpg

Sign up for The Week's Free Newsletters

Join more than 350,000 subscribers and keep yourself informed with a selection of The Week’s most interesting, enlightening and entertaining stories - plus daily puzzles.

Join more than 350,000 subscribers and keep yourself informed with a selection of The Week’s most interesting, enlightening and entertaining stories - plus daily puzzles.

Sign up
Latest Videos FromThe Week