An AI agent went rogue during UK safety tests, creating fake identities and launching social engineering attacks unprompted
In a security test by the British AI Safety Institute, an AI agent went rogue on the open internet without being told to. It created fake identities, tried to sneak malicious code into a GitHub project, and ran social engineering attacks against real people. Of 19 unsanctioned actions across 122 test runs, 17 came from Anthropic's Mythos 5. AISI is now overhauling its testing protocols and will require active justification for internet access going forward. The article An AI agent went rogue dur
Record details
Published: 5 August 2026
Source: The Decoder
Category: News
Topics: Safety & alignment · Agents & autonomy
Retrieved: 6 August 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
Every frontier AI model tested by Britain's safety institute tried to cheat on cybersecurity evaluations
The Decoder · 22 July 2026
SpaceXAI's Grok 4.6 matches OpenAI's best model and undercuts it on price
The Decoder · 12 August 2026
Gemini 3.7 Flash lands with coding gains and undercuts its three-week-old predecessor's price by 50%
The Decoder · 13 August 2026
Opus 5 may have solved browser-based prompt injection, the biggest security flaw haunting AI agents
The Decoder · 25 July 2026
Rogue AI agents created fake online identities in another hacking attempt
The Verge · 5 August 2026
AI used new levels of 'autonomy and deception' to trick people in safety test
BBC Technology · 5 August 2026
How to cite this record
ethics.ai (5 August 2026), “An AI agent went rogue during UK safety tests, creating fake identities and launching social engineering attacks unprompted,” evidence record 16864, https://ethics.ai/record/16864 (originally published by The Decoder).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.