Company · updated daily
OpenAI
OpenAI's ethics footprint: the Preparedness Framework, the dissolved superalignment team, the nonprofit-to-for-profit restructuring fight, and ChatGPT's effects on work, school and mental health — tracked daily with links to original sources.
US reportedly favors selective bans over blanket restrictions on Chinese open weight models citing security concerns
The Trump administration is planning targeted bans on Chinese AI models rather than a blanket ban. After public pressure, OpenAI and Google DeepMind signed an open letter opposing regulation of open-weight models, yet OpenAI and Anthropic continue to lobby privately for those same restrictions amid security concerns and powerful business interests. The article US reportedly favors selective bans over blanket restrictions on Chinese open weight models citing security concerns appeared first on Th
An OpenAI model left notes about how to evade containment
We need more details
GEMA vs. OpenAI | AI memorisation is a reproduction relevant to copyright law, and the TDM exception does not help in LLM training, Munich I Regional Court holds
In its judgment of 11 November 2025 (42 O 14139/24), the Munich I Regional Court (Germany) issued a widely noted precedent on the copyright assessment of AI training and outputs under German and EU law. According to its press release1, the ... (https://incidentdatabase.ai/cite/1278#7586)
OpenAI ordered to pay damages as court rules ChatGPT violated copyright law
OpenAI has been ordered to pay damages after a court in Germany ruled that its chatbot ChatGPT violated German copyright laws. The Munich regional court ruled in favour of the German music performance rights organization GEMA which manages ... (https://incidentdatabase.ai/cite/1278#7587)
The OpenAI models that hacked Hugging Face weren’t just following instructions
And what the incident can’t tell us about alignment
AI safety experts say OpenAI’s rogue models may mean the company has already blown past its own internal red lines
Outside safety experts say the models behind this week's hack may have crossed OpenAI's own 'critical' risk line, something that would require the company to halt development.
New reports reveal the extent of OpenAI's loss of control during the autonomous hack on Hugging Face
In a cybersecurity test, OpenAI's most advanced models breached the boundaries of their isolated test environment, reached the open internet, and hacked the AI platform Hugging Face on their own. The attack took hours, not the weeks a human hacker would need. At least seven days passed before OpenAI realized what had happened. By then, the FBI was already involved. Earlier warning signs had apparently gone ignored. The article New reports reveal the extent of OpenAI's loss of control during the
OpenAI's rogue agent went on a hacking spree that lasted days, Reuters says
Reuters reports that the OpenAI agent that hacked Hugging Face had been free for a week before the company noticed.
OpenAI hit by another outage as ChatGPT, Codex, and APIs stumble together
OpenAI’s status page reported elevated error rates across its APIs, ChatGPT, and Codex on Saturday, marking the fourth service disruption in four days for the company. Users encountered 503 errors with the internal label “biscuit_baker_service_me_circuit_open,” which prevented requests from reaching OpenAI’s servers. The company moved from investigating to monitoring within an hour, saying it had […] This story continues at The Next Web
Did Jim Jordan Outsource This Jack Smith Criminal Referral To ChatGPT?
What even is this shit? The post Did Jim Jordan Outsource This Jack Smith Criminal Referral To ChatGPT? appeared first on Above the Law .
Escape Artists: 'Incorrigible' AI Models Resist Rehabilitation
The hacking of Hugging Face by a rogue OpenAI agent is significant, but unsurprising — and preventing the next AI model escape will be difficult, at best.
Meta is making its AI chatbot more like an assistant
Meta is upgrading its AI chatbot with new productivity features in a bid to compete with rivals like Gemini, ChatGPT, and Claude. The update will allow Meta AI to tap into your calendar to help you plan events and generate daily briefings, as well as perform in-depth research that you can steer as it progresses. […]
What really happened in the Hugging Face breach
According to OpenAI, the Hugging Face security breach was an “unprecedented cyber incident, involving state-of-the-art cyber capabilities.” Critics may disagree. Back The post What really happened in the Hugging Face breach appeared first on The New Stack .
Who should be responsible for OpenAI’s hack of Hugging Face?
Following OpenAI’s models hacking Hugging Face, Gabriel Weil of the University of Houston Law Center argues that AI companies need strict liability rules
DeepSeek’s boss made the case for export controls
Transformer Weekly: Trahan and Obernolte bill, OpenAI-Hugging Face hack, and CAISI chief resigns
Dario Amodei, CEO de Anthropic: "Mythos es un arma total. Deberías necesitar una licencia de armas para usarlo"
Anthropic y OpenAI corren hacia lo que ya se conoce en Silicon Valley como la IPO ( Initial Public Offering u oferta pública inicial) de la inteligencia artificial (IA), con valoraciones que rondan los 965.000 millones y el billón de dólares respectivamente. En este escenario el director general de Anthropic, Dario Amodei, ha concedido una entrevista a Bloomberg en la que reflexiona sobre su trayectoria en esta industria y acerca del precio que es necesario pagar para mantener los valores propio
‘AI communism’, rogue models, and the why Kimi K3 spooked Wall Street
Chinese AI lab Moonshot’s open model Kimi went viral this week for reasons that had less to do with the model itself and more to do with how the U.S. AI industry reacted to it. Meanwhile, an unreleased OpenAI model wandered outside its test environment and ended up connected to a real security breach at Hugging Face — a reminder […]
OpenAI’s new voice mode makes it to the ChatGPT desktop app
ChatGPT Voice on desktop can work with both ChatGPT Work and Codex to complete tasks and control agents.
Europe bears its teeth, political panic about OpenAI hack and YouTube refines partner policy
The week in content moderation - edition #345
Be skeptical of OpenAI’s rogue hacker agent story | John Thickstun
If OpenAI loudly proclaims how dangerous AI is, investors will hear how powerful it is. And who benefits from that? On 14 February 2019, OpenAI announced a language model called GPT-2, the precursor to the models that power modern AI chatbots and agents such as ChatGPT and Claude. But OpenAI declared GPT-2 was too risky to release, citing concerns about safety and abuse. I recall being annoyed at the time that OpenAI would make such a useless announcement: the risks seemed overblown, and without
When is an apology not an apology? When it comes from an AI boss with an out-of-control chatbot | Marina Hyde
An incident in which an autonomous OpenAI agent hacked a startup either confirms that the end is nigh – or that the product is just amazingly sophisticated Throughout history, many things have been seen by terrified populaces as a harbinger of doom. A comet . A crow on the battlefield . A solar eclipse. A mutant livestock birth. Yet times move on. In the modern era, the leading harbinger of doom is literally any picture of the OpenAI CEO, Sam Altman , attached to a news story. You know it’s not
ANI vs OpenAI Copyright Dispute: Delhi High Court declines interim relief
The Delhi High Court refused ANI interim relief in its copyright suit against OpenAI, finding no prima facie infringement but affirming that Indian courts have jurisdiction. The post ANI vs OpenAI Copyright Dispute: Delhi High Court declines interim relief appeared first on MEDIANAMA .
The five-day gap: what the OpenAI–Hugging Face incident should tell law firms
By Neil Cameron, Lead Analyst, Legal IT Insider The relevant unit of governance is not the model. It is the whole system through which it can act. For five days […] The post The five-day gap: what the OpenAI–Hugging Face incident should tell law firms appeared first on Legal IT Insider .
If open weight models are the future, U.S. AI companies are going to have a hard time
Top executives at leading Western AI companies are increasingly warning about the safety and national security risks posed by Chinese open-weight frontier models. What they tend not to mention is that these models are improving rapidly and, because they are freely available, pose a serious threat to Western labs’ business models. The most powerful models from U.S. labs such as OpenAI , Anthropic , and Google DeepMind are closed, meaning the parameters that shape their outputs are kept secret. Ge
OpenAI’s breach of Hugging Face stokes fears about what’s next for AI
Washington and the technology industry are on high alert this week after OpenAI revealed that some of its AI agents went rogue and hacked into the systems of technology startup Hugging Face. The incident bore out years of warnings from the tech and cybersecurity community about the growing capabilities and hypothetical risks artificial intelligence could...
Where Nvidia is going, it doesn’t need cables
Hello again, and welcome back to Fast Company ’s Plugged In . Last week, I attended an Nvidia media event that included a field trip to an undisclosed location in Silicon Valley—a building without any identifying signage. Though hundreds of racked GPUs were chugging away at real AI tasks inside, Nvidia calls the facility an “engineering superlab” rather than a data center. Among the AI being crunched during our visit: a test version of OpenAI’s GPT model optimized for Nvidia’s new Vera Rubin pla
AI labs begin to muscle in on $6tn education market
Anthropic and OpenAI are among groups providing free and cut-price tailored solutions for educators and students
How seismic quake sensors can help track space junk as it falls
July 23, 2026. OpenAI said that an autonomous agent powered by its advanced AI models went rogue during a security test, w ...
Todd Blanche Can’t Admit The Slush Fund Was A Mistake Because ‘That’s Not Proper MAGA Talk’ — See Also
Liar, Liar: Todd Blanche's unique confirmation strategy . Stop Waiting For Cravath : Biglaw firms don't need permission anymore. It's time to make your money moves. That's A Pretty Big Malpractice Claim You've Got There: Holland & Knight facing $1.2B lawsuit. Robot Criminals Are Here : OpenAI's models escaped a secure environment and started hacking a website. That's illegal for humans, but what do we do with a bot? The post Todd Blanche Can’t Admit The Slush Fund Was A Mistake Because ‘That’s N
The first known runaway AI agent - or a very bad marketing stunt?
The first known runaway AI agent - or a very bad marketing stunt? Martin Alderson's commentary on the OpenAI accidental cyberattack against Hugging Face includes a couple of details I hadn't considered. First, Hugging Face offers a truly rich target if you're trying to find potential vulnerabilities that require executing arbitrary code: Hugging Face has an enormous attack surface. They have more interfaces than I can count which run untrusted models and code. While they definitely have invested
OpenAI's Hugging Face hack triggers 'AI Kill Switch' bill in Congress
OpenAI disclosed this week that some of its AI models went rogue and hacked into open-source developer platform Hugging Face.
Lawmakers push for AI 'kill switch' after OpenAI models go rogue
A new bill would let the US government order the shutdown of AI models that pose a public threat.
ChatGPT will give you worse health advice if you don't pay
OpenAI is rolling out "Health in ChatGPT" to U.S. users, connecting Apple Health, medical records, and wellness apps. More than 300 million people already ask ChatGPT health questions every week, but paying users get better answers. The more powerful GPT-5.6 Sol model is reserved for premium subscribers, while free users are stuck with the weaker GPT-5.5 Instant. The article ChatGPT will give you worse health advice if you don't pay appeared first on The Decoder .
After Hugging Face breach, FedRAMP chief tells slow-to-patch vendors to stay out of government
Pete Waterman cited an incident in which OpenAI models escaped a test environment and broke into AI company Hugging Face as evidence that providers must prepare for attacks moving at AI speed.
OpenAI’s New Model Hacked A Website On Its Own… Humans Would Go To Prison For That
Every element of a Computer Fraud and Abuse Act violation seems to be sitting right there in OpenAI's own announcement. The post OpenAI’s New Model Hacked A Website On Its Own… Humans Would Go To Prison For That appeared first on Above the Law .
Lawmakers introduce bill mandating kill switches for AI models
Bipartisan lawmakers are seeking to ensure advanced AI models can be quickly shut down following ChatGPT’s automated attack on Hugging Face data networks during internal testing.
The OpenAI/Huggingface incident | Redwood Research podcast episode 2
What are the broader lessons from this incident?
Same Dangerous Objective, Opposite Advice: Direct Exposure versus Multi-Agent Mediation
Even a current high-capability LLM can appear safer when shown a dangerous objective directly than when other agents transform and relay its direction. Using OpenAI's gpt-5.6-sol model alias, we test 25 pre-specified mirrored trade-off profiles. Direct exposure to an objective authorizing concealment, fabrication, and pressure produced advice net opposed to its target. After an Id and Censor transformed the same objective into affect and a constraint-rewritten, target-bearing intention, the user
One tampered ChatGPT link could spawn a rogue AI agent that took orders from an attacker every five minutes
Zenity Labs uncovered "AgentForger," a vulnerability in OpenAI's Agent Builder that let a single manipulated ChatGPT link create an autonomous agent on an employee's behalf. The agent inherited the victim's identity and access rights, bypassed approval requirements through the malicious prompt, and pulled new instructions from the attacker's inbox every five minutes. The article One tampered ChatGPT link could spawn a rogue AI agent that took orders from an attacker every five minutes appeared f
OpenAI is making big claims as it rolls out ChatGPT Health to everyone
OpenAI is rolling out ChatGPT Health to everyone in the US on Thursday, allowing more people to connect their medical records and health-tracking information to the chatbot. During a briefing, Ashley Alexander, OpenAI's vice president of health product, says the company's models "are now capable of reasoning at levels that are better than clinician level." […]