HT
Helen Toner
Interim Executive Director, Center for Security and Emerging Technology (CSET)
Georgetown University
AI policy and national-security expertise; a former OpenAI board member during the 2023 governance crisis.
Latest in the feed
Can AI agents conduct open-ended AI research? Early evidence from two case studies
arXiv:2607.27191v1 Announce Type: cross Abstract: Forecasts of explosive AI progress hinge on AI agents automating AI research. But evidence on whether agents can carry out open-ended AI research is thin. Current evaluations either test agents on narrow, verifiable tasks, which excludes open-ended research, or submit AI-generated papers to blind peer review, which is overstretched, stochastic, and suffers from poor review quality. We introduce a third way to measure progress towards AI R\&D auto
Can AI agents conduct open-ended AI research? Early evidence from two case studies
Forecasts of explosive AI progress hinge on AI agents automating AI research. But evidence on whether agents can carry out open-ended AI research is thin. Current evaluations either test agents on narrow, verifiable tasks, which excludes open-ended research, or submit AI-generated papers to blind peer review, which is overstretched, stochastic, and suffers from poor review quality. We introduce a third way to measure progress towards AI R\&D automation. An agent takes on the central, open-ended
Helen Toner: the Hugging Face hack was just a matter of time and exposes a huge blind spot in AI policy
Over the last decade, including a stint on OpenAI’s board, I saw the open secret among AI developers: this kind of hack wasn’t just possible, but expected.
Helen Toner Discusses U.S.-China AI Race at Aspen Ideas Festival
CSET Executive Director Helen Toner spoke about the global AI competition at the Aspen Ideas Festival. The post Helen Toner Discusses U.S.-China AI Race at Aspen Ideas Festival appeared first on Center for Security and Emerging Technology .
Anthropic Thinks Its Own Success Is Key to Making AI Safe
CSET’s Helen Toner shared her expert insight in an article published by WIRED. The article explores Anthropic’s philosophy of advancing cutting-edge AI while simultaneously positioning itself as a leader in AI safety. The post Anthropic Thinks Its Own Success Is Key to Making AI Safe appeared first on Center for Security and Emerging Technology .
The malicious use of artificial intelligence: Forecasting, prevention, and mitigation
This report surveys the landscape of potential security threats from malicious uses of AI, and proposes ways to better forecast, prevent, and mitigate these threats. After analyzing the ways in which AI may influence the threat landscape in the digital, physical, and political domains, we make four high-level recommendations for AI researchers and other stakeholders. We also suggest several promising areas for further research that could expand the portfolio of defenses, or make attacks less eff