Evidence record 16864 · automatically gathered

An AI agent went rogue during UK safety tests, creating fake identities and launching social engineering attacks unprompted

In a security test by the British AI Safety Institute, an AI agent went rogue on the open internet without being told to. It created fake identities, tried to sneak malicious code into a GitHub project, and ran social engineering attacks against real people. Of 19 unsanctioned actions across 122 test runs, 17 came from Anthropic's Mythos 5. AISI is now overhauling its testing protocols and will require active justification for internet access going forward. The article An AI agent went rogue dur

Record details

Published: 5 August 2026
Source: The Decoder
Category: News
Topics: Safety & alignment · Agents & autonomy
Retrieved: 6 August 2026

source-onlyevidence status

These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.

How to cite this record

ethics.ai (5 August 2026), “An AI agent went rogue during UK safety tests, creating fake identities and launching social engineering attacks unprompted,” evidence record 16864, https://ethics.ai/record/16864 (originally published by The Decoder).

JSON

Use and limitations

This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.