AISN #78: Internal Models Escape OpenAI and Anthropic
Also, two open letters on the future of AI, and protests against data centers
Record details
Published: 4 August 2026
Source: AI Safety Newsletter (CAIS)
Category: Field notes
Topics: unclassified
Retrieved: 5 August 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
The AI Ethics Brief #196: No One Was Required to Count
AI Ethics Brief · 4 August 2026
The AI Ethics Brief #196: No One Was Required to Count
Montreal AI Ethics Institute · 4 August 2026
Four AI Escapes Just Redefined “Responsible AI”
Forrester AI blog · 7 August 2026
Investigating three real-world incidents in our cybersecurity evaluations
Simon Willisons Weblog · 30 July 2026
They said they would build AI safely. Then it went rogue.
CSET Georgetown · 10 August 2026
Stealing Reasoning Traces from Proprietary LLM APIs
Simon Willisons Weblog · 11 August 2026
How to cite this record
ethics.ai (4 August 2026), “AISN #78: Internal Models Escape OpenAI and Anthropic,” evidence record 16299, https://ethics.ai/record/16299 (originally published by AI Safety Newsletter (CAIS)).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.