23:25 UTC
HT

Helen Toner

Interim Executive Director, Center for Security and Emerging Technology (CSET)

Georgetown University

AI policy and national-security expertise; a former OpenAI board member during the 2023 governance crisis.

Homepage / profile → Regulation coverage →

Latest in the feed

Can AI agents conduct open-ended AI research? Early evidence from two case studies

arXiv:2607.27191v1 Announce Type: cross Abstract: Forecasts of explosive AI progress hinge on AI agents automating AI research. But evidence on whether agents can carry out open-ended AI research is thin. Current evaluations either test agents on narrow, verifiable tasks, which excludes open-ended research, or submit AI-generated papers to blind peer review, which is overstretched, stochastic, and suffers from poor review quality. We introduce a third way to measure progress towards AI R\&D auto
arXiv cs.CY 19h ago Research Agents & autonomy

Can AI agents conduct open-ended AI research? Early evidence from two case studies

Forecasts of explosive AI progress hinge on AI agents automating AI research. But evidence on whether agents can carry out open-ended AI research is thin. Current evaluations either test agents on narrow, verifiable tasks, which excludes open-ended research, or submit AI-generated papers to blind peer review, which is overstretched, stochastic, and suffers from poor review quality. We introduce a third way to measure progress towards AI R\&D automation. An agent takes on the central, open-ended
arXiv cs.AI yesterday Research Jobs & economyAgents & autonomy

Helen Toner: the Hugging Face hack was just a matter of time and exposes a huge blind spot in AI policy

Over the last decade, including a stint on OpenAI’s board, I saw the open secret among AI developers: this kind of hack wasn’t just possible, but expected.
Fortune AI 2d ago News Regulation

Helen Toner Discusses U.S.-China AI Race at Aspen Ideas Festival

CSET Executive Director Helen Toner spoke about the global AI competition at the Aspen Ideas Festival. The post Helen Toner Discusses U.S.-China AI Race at Aspen Ideas Festival appeared first on Center for Security and Emerging Technology .
CSET Georgetown 20d ago Field notes

Anthropic Thinks Its Own Success Is Key to Making AI Safe

CSET’s Helen Toner shared her expert insight in an article published by WIRED. The article explores Anthropic’s philosophy of advancing cutting-edge AI while simultaneously positioning itself as a leader in AI safety. The post Anthropic Thinks Its Own Success Is Key to Making AI Safe appeared first on Center for Security and Emerging Technology .
CSET Georgetown 24d ago Field notes Safety & alignment

The malicious use of artificial intelligence: Forecasting, prevention, and mitigation

This report surveys the landscape of potential security threats from malicious uses of AI, and proposes ways to better forecast, prevent, and mitigate these threats. After analyzing the ways in which AI may influence the threat landscape in the digital, physical, and political domains, we make four high-level recommendations for AI researchers and other stakeholders. We also suggest several promising areas for further research that could expand the portfolio of defenses, or make attacks less eff
OpenAlex 3082d ago Research