Evidence record 10762 · automatically gathered

OpenAI built an AI super-hacker to break its own models, then locked it away

OpenAI has trained an elite hacker, then locked it in a cage. Its whole job is to break OpenAI’s own AI. The company says it is too dangerous to let anyone else near it. The model is called GPT-Red, and OpenAI detailed it this week. It is an automated red-teamer: software that hunts for ways […] This story continues at The Next Web

Record details

Published: 15 July 2026
Source: The Next Web AI
Category: News
Topics: Safety & alignment · Jobs & economy
Retrieved: 16 July 2026

source-onlyevidence status

These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.

How to cite this record

ethics.ai (15 July 2026), “OpenAI built an AI super-hacker to break its own models, then locked it away,” evidence record 10762, https://ethics.ai/record/10762 (originally published by The Next Web AI).

JSON

Use and limitations

This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.