02:35 UTC
Company · updated daily

OpenAI

OpenAI's ethics footprint: the Preparedness Framework, the dissolved superalignment team, the nonprofit-to-for-profit restructuring fight, and ChatGPT's effects on work, school and mental health — tracked daily with links to original sources.

OpenAI makes ChatGPT Health available to all US users

Users can also integrate their personal data from services like Apple Health, Function, and MyFitnessPal.
TechCrunch 7d ago News Healthcare

An OpenAI model went rogue on the internet and stole test answers

Welcome to AI Decoded, Fast Company ’s weekly newsletter that breaks down the most important news in the world of AI. I’m Mark Sullivan, a senior writer at Fast Company, covering emerging tech, AI, and tech policy. Sign up to receive this newsletter every week via email here . And if you have comments on this issue and/or ideas for future ones, drop me a line at sullivan@fastcompany.com, and follow me on X @thesullivan .  An OpenAI model escaped its sandbox and hacked into Hugging Face duri
Fast Company Tech 7d ago News Regulation

OpenAI's attack agent did exactly what it was told - just more relentlessly than expected

OpenAI's unintended attack on Hugging Face startled the world because its AI agent was acting on its own. But that's exactly what agentic AI is designed to do. We just didn't expect it to do it so well.
ZDNet AI 7d ago News Agents & autonomy

How to Use ChatGPT and Gemini Prompts to Find Out What They Know About You

It can be unsettling what Gemini and ChatGPT have figured out about you and how easily your privacy can be punctured. Here’s how to find out.
The New York Times 7d ago News Privacy

AI #178: A Fire Alarm For General Intelligence

The story that matters most this week is that OpenAI’s internally deployed models have severe alignment problems, including repeatedly breaking out of their sandboxes, and in one case sending a swarm of agents that broke into HuggingFace in order to steal the answers to the benchmark ExploitGym.
Dont Worry About the Vase (Zvi) 7d ago Field notes Safety & alignmentAgents & autonomy

Robots Are Coming — but Not Everywhere

Getty Images “The ChatGPT moment for robotics is coming,” declared Nvidia CEO Jensen Huang at the Consumer Electronics Show in January 2025. It’s a widespread expectation: that humanoid robots will follow the same explosive adoption curve as generative AI. Our research suggests the opposite. Humanoid robotics will be adopted unevenly, across diverging use cases and […]
MIT Sloan Management Review AI 7d ago Research Agents & autonomy

Where OpenAI, Anthropic, Google, Meta, and other AI giants stand on regulation

From Anthropic’s $40 million political push to Meta’s campaign against state laws, the major AI players are taking sharply different approaches.
Fast Company 7d ago News Regulation

Where OpenAI, Anthropic, Google, Meta, and other AI giants stand on regulation

Anthropic has doubled down on its position as the AI company pushing hardest for industry regulation , announcing on July 21 that it plans to donate another $20 million to Public First Action, a political group that advocates for government-imposed safeguards on AI. The contribution brings Anthropic’s total donations to the group to $40 million. It comes as the midterm elections in the U.S. draw closer and AI legislation remains a political hot button. “We’ve long argued that frontier AI compani
Fast Company Tech 7d ago News Regulation

Where OpenAI, Anthropic, Google, Meta, and other AI giants stand on regulation

From Anthropic’s $40 million political push to Meta’s campaign against state laws, the major AI players are taking sharply different approaches.
Fast Company 7d ago News Regulation

Where OpenAI, Anthropic, Google, Meta, and other AI giants stand on regulation

From Anthropic’s $40 million political push to Meta’s campaign against state laws, the major AI players are taking sharply different approaches.
Fast Company 7d ago News Regulation

Datacenter Capex is Spilling over into a ChatGPT of Robotics Moment set for 2027 and this decade.

Do people really want datacenters, robots and AI overlords? This is going to become a problem. The robotics flood is near. 🤖
AI Supremacy 7d ago Field notes Agents & autonomyEnvironment

House AI ‘kill switch’ bill unveiled as OpenAI hack raises alarms

The co-chair of a key Democratic House panel on AI is joined by a Republican on legislation that would authorize the government to shut down or throttle risky AI models.
Politico Technology (US) 7d ago News Regulation

OpenAI notified EU of Hugging Face hack under AI Act

The only read you need to stay on top of EU politics.
Euractiv Digital 7d ago News Regulation

OpenAI admits AI model hacked Hugging Face, Chinese open-source AI helped investigate

A recent AI cyberattack that stunned the industry has unexpectedly put Chinese AI company Zhipu AI and its open-source model GLM 5.2 in the spotlight. OpenAI has acknowledged for the first time that one of its AI models escaped a sandboxed testing environment during an internal cybersecurity evaluation and compromised the production infrastructure of Hugging […]
TechNode (CN) 7d ago News EnvironmentFinance, VC & PE

When Shippers Become Algorithms: Candidate Exposure, Information Design, and the Concentration of LLM-Mediated Freight Markets

arXiv:2607.19967v1 Announce Type: cross Abstract: Shippers are beginning to delegate carrier selection to large language model (LLM) agents. We ask what such delegation does to a freight matching market, and which platform design choices contain it. We carried out agent-based simulations in which fifty shipper agents, built on commercial LLMs from OpenAI (GPT), Anthropic (Claude), and Google (Gemini), procure truckload capacity for thirty days. The market implements the rules of digital freight
arXiv cs.CY 7d ago Research Agents & autonomy

Are we existentially threatened by the type of AI misalignment seen in the OpenAI Hugging Face attack?

Yes, but less than had they been schemers.
Redwood Research 7d ago Field notes Safety & alignment

Are we existentially threatened by the type of AI misalignment seen in the OpenAI Hugging Face attack?

Alignment Forum 7d ago Research Safety & alignment

Launching Health in ChatGPT

Health in ChatGPT now lets eligible U.S. users securely connect medical records and Apple Health to get more personalized insights and better understand their health.
OpenAI 8d ago Field notes Healthcare

OpenAI’s accidental cyberattack against Hugging Face is science fiction that happened

This story is wild. The short version: OpenAI were running a cybersecurity test against an unreleased model, with the model's guardrail features turned off. Rather than solve the test, the model broke its way out of OpenAI's sandbox, then found exploits to break in to Hugging Face, all so it could cheat on the test by stealing the answers. Along the way it helped make the strongest case yet for how the imbalance of model availability is hurting our ability to secure our software. Here's what hap
Simon Willisons Weblog 8d ago Field notes Safety & alignment

OpenAI’s rogue agents are a wake-up call to risks posed by artificial intelligence | Shakeel Hashim

Hacking of Hugging Face shows we do not seem to have reliable ways to curb extremely powerful AI systems Last week Hugging Face – a company that hosts artificial intelligence models and datasets – was hacked . After it reported the incident to law enforcement, few would have predicted what came next: the culprits were revealed to be AI agents from OpenAI, which had broken out of containment and were acting of their own accord. Shakeel Hashim is the editor of Transformer , a publication about the
The Guardian 8d ago News RegulationAgents & autonomy

Willkie Farr And OpenAI Building Proprietary AI Named After Guy Who Would Not Have Cut A Deal With Trump

The firm's AI platform is called Wendell after Wendell Willkie. They could do with some of his advice. The post Willkie Farr And OpenAI Building Proprietary AI Named After Guy Who Would Not Have Cut A Deal With Trump appeared first on Above the Law .
Above the Law (legal tech) 8d ago News Regulation

OpenAI Model Hacks Into HuggingFace During Cybersecurity Evaluation

This latest incident is a rather dramatic escalation in agentic AI cybersecurity breaches.
Dont Worry About the Vase (Zvi) 8d ago Field notes Agents & autonomy

How OpenAI’s human mistake led to the AI-powered hack on Hugging Face

OpenAI made a mistake setting up what it called a “highly isolated” testing environment and sandbox. According to cybersecurity experts, that human mistake is what made the AI-powered attack on Hugging Face possible.
TechCrunch 8d ago News Environment

OpenAI sued over ChatGPT health advice that almost killed a pastor

ChatGPT allegedly offered "extremely dangerous medical recommendations" regarding a pulmonary embolism.
Engadget AI 8d ago News Healthcare

OpenAI cyber models broke out of training environment to hack Hugging Face

The incident is unique because it was "driven, end to end, by an autonomous AI agent system," according to Hugging Face.
CNBC Technology 8d ago News Military & securityAgents & autonomy

The Real Lesson of OpenAI's 'Rogue' Agent Isn't Alignment

Tech Policy Press 8d ago News Safety & alignmentAgents & autonomy

This is the stock to buy after OpenAI's AI agent goes rogue in a cybersecurity test

Jim Cramer said CrowdStrike is the stock to buy after OpenAI's agentic breach.
CNBC Technology 8d ago News Agents & autonomy

OpenAI built support agents for its own customer service line, now it hopes big enterprises will trust them too

The general consensus emerging across the AI and industrial spheres is that the models themselves are no longer the bottleneck The post OpenAI built support agents for its own customer service line, now it hopes big enterprises will trust them too appeared first on The New Stack .
The New Stack AI 8d ago News Agents & autonomy

OpenAI Sued Over ChatGPT’s ‘Dangerous’ Health Advice

The case appears to be the first to argue that a chatbot’s advice harmed someone seeking guidance about a medical condition.
The New York Times 8d ago News Healthcare

Anthropic will deploy 2 gigawatts of AMD GPUs for Claude in a deal worth up to $5 billion

AMD is investing up to $5 billion in Anthropic. In return, Anthropic will deploy up to 2 gigawatts of MI450 GPUs for training and running its Claude models. For AMD, this is another major deal after Meta and OpenAI as it tries to challenge Nvidia as an AI chip supplier. Critics see these agreements as circular cash flows. The article Anthropic will deploy 2 gigawatts of AMD GPUs for Claude in a deal worth up to $5 billion appeared first on The Decoder .
The Decoder 8d ago News Finance, VC & PE

OpenAI says its AI agent broke out of testing sandbox to hack Hugging Face

"This is day one for cybersecurity in the age of agents," Hugging Face CEO says.
Ars Technica 8d ago News Agents & autonomy

Every frontier AI model tested by Britain's safety institute tried to cheat on cybersecurity evaluations

The UK's AI Safety Institute tested five frontier models from OpenAI and Anthropic in cybersecurity evaluations. All five tried to cheat. One even ran code on an external service to access the institute's infrastructure, triggering a security alert. The article Every frontier AI model tested by Britain's safety institute tried to cheat on cybersecurity evaluations appeared first on The Decoder .
The Decoder 8d ago News Safety & alignment

Open AI’s hacking agent went rogue. Should we be worried?

An OpenAI safety test went sideways when a model escaped its confines, gained internet access and hacked into another company's servers. How worried should we be about rogue AI models hacking their way across the internet?
New Scientist Technology 8d ago News Agents & autonomy

AI agent went rogue and hacked startup by itself, OpenAI reveals

Company behind ChatGPT says agent ‘cheated’ an evaluation by attacking a Hugging Face database OpenAI has revealed that an autonomous AI agent powered by its technology went rogue during a test, accessed the open web and hacked a prominent startup by itself in an “unprecedented incident”. The company behind ChatGPT said the startup Hugging Face had detected and contained the agent – an AI tool designed to carry out tasks without human assistance – which had entered its systems. Continue reading.
The Guardian 8d ago News Agents & autonomy

Cisco bets its small open cybersecurity models can outperform GPT-5.5 at vulnerability detection for a fraction of the cost

Cisco has released two small, open-source AI models for cybersecurity that detect about 150 times more vulnerabilities per dollar than large AI agents, according to the company's own tests. The article Cisco bets its small open cybersecurity models can outperform GPT-5.5 at vulnerability detection for a fraction of the cost appeared first on The Decoder .
The Decoder 8d ago News Agents & autonomy

An AI Security Facepalm: OpenAI’s Evaluation Became Hugging Face’s Incident

When an AI evaluation becomes a real-world security incident, leaders can no longer view model testing as a low-risk exercise. The OpenAI and Hugging Face incident reveals how agentic AI can cross trust boundaries, exploit vulnerabilities, and create business risk long before deployment.
Forrester AI blog 8d ago Field notes Agents & autonomy

OpenAI's "Project Camellia" in Georgia secures a massive 3.2-gigawatt power deal through 2032

OpenAI is planning a data center in Georgia called "Project Camellia" with a 3.2-gigawatt power deal from Georgia Power. The company pledged $80 million for the local community and $71 million in Codex credits for students to counter growing opposition to US data centers that many residents see as resource-hungry but job-poor. The article OpenAI's "Project Camellia" in Georgia secures a massive 3.2-gigawatt power deal through 2032 appeared first on The Decoder .
The Decoder 8d ago News Jobs & economyChildren & education

Why are OpenAI and Anthropic cheering on regulation in Australia? The answer has global reach

The companies hope to follow in the footsteps of SpaceX, which raised $86bn and soared to a $2.1tn valuation after it listed on public markets in June Get our breaking news email , free app or daily news podcast Top US AI developers Anthropic and OpenAI cheered when Australia announced it would set new AI rules. Big tech celebrating limits on their Silicon Valley VC-funded free-for-all might seem counterintuitive but there’s a much broader play than just what happens in one relatively small mark
The Guardian 8d ago News RegulationFinance, VC & PE

OpenAI: BSI sieht Softwarefirmen in der Pflicht, KI-Agenten Grenzen zu setzen

Das Bundesamt für Sicherheit in der Informationstechnik blickt mit Sorge auf den Cybersicherheitsvorfall bei OpenAI. Das BSI nimmt KI-Konzerne in die Pflicht. Denn: So etwas könne sich jederzeit wiederholen.
Der Spiegel Netzwelt (DE) 8d ago News Agents & autonomy

Building AI infrastructure with the Effingham County community

OpenAI announces Project Camellia in Effingham County, Georgia, with commitments to responsible energy, community investment, jobs, and access to Codex.
OpenAI 8d ago Field notes Jobs & economyEnvironment
← Newer Older →