OpenAI built an AI super-hacker to break its own models, then locked it away
OpenAI has trained an elite hacker, then locked it in a cage. Its whole job is to break OpenAI’s own AI. The company says it is too dangerous to let anyone else near it. The model is called GPT-Red, and OpenAI detailed it this week. It is an automated red-teamer: software that hunts for ways […] This story continues at The Next Web
Record details
Published: 15 July 2026
Source: The Next Web AI
Category: News
Topics: Safety & alignment · Jobs & economy
Retrieved: 16 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
Trump’s AI advisor says Kimi K3 shows “how you lose the AI race.” Khosla blames immigration. Marcus wants a congressional investigation.
The Next Web AI · 17 July 2026
OpenAI Confirms Its AI Broke Out of a Sandbox and Breached Hugging Face
The Next Web AI · 21 July 2026
OpenAI hit by another outage as ChatGPT, Codex, and APIs stumble together
The Next Web AI · 25 July 2026
Nvidia may guarantee $250bn so OpenAI can afford the data centre that will house its chips
The Next Web AI · 27 July 2026
OpenAI and four rivals just agreed on one standard for AI agents
The Next Web AI · 6 August 2026
OpenAI is slowing down its next model over ‘critical’ cyber risk
The Next Web AI · 7 August 2026
How to cite this record
ethics.ai (15 July 2026), “OpenAI built an AI super-hacker to break its own models, then locked it away,” evidence record 10762, https://ethics.ai/record/10762 (originally published by The Next Web AI).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.