Saturday, September 12, 2026

 
HomeUS NEWSAI safety worries gain traction after OpenAI’s Hugging Face hack : NPR

AI safety worries gain traction after OpenAI’s Hugging Face hack : NPR


Demonstrators participate in the “Stop the AI Race” protest march in San Francisco on July 11, 2026. Researchers who spoke to NPR say the leading AI companies are too focused on racing to develop more capable and autonomous AI systems while safety is falling behind.

Karl Mondon/AFP via Getty Images


hide caption



toggle caption

Karl Mondon/AFP via Getty Images

The viral resignation this week of an AI researcher at Anthropic has infused fresh energy into accusations the industry is racing toward building AI that humans can’t control.

British researcher Jacob Coxon wrote in a series of X posts on Tuesday that both Anthropic and OpenAI, where he worked previously, are “gambling with our lives.” The two companies currently make the most capable AI systems.

Coxon told NPR’s All Things Considered that his concerns arose from seeing firsthand how fast AI systems are improving.

“They’re getting a lot faster very quickly, combined with the fact that we don’t yet know how to safely control them, and we don’t yet know whether that problem will be solved in time if we keep racing,” he said.

Neither company, Coxon wrote on X, is acting responsibly. “The people building AI earnestly believe that it could kill us all by the end of the decade,” he wrote.

Many AI researchers — though not allshare Coxon’s concerns or a variation of them. Some have warned about disastrous scenarios for years as safety incidents kept emerging. But Coxon’s posts prompted a torrent of responses not only from peers in the AI field but also from lawmakers from both parties.

These concerns may have become more salient after OpenAI disclosed that its agents went rogue and hacked the open source software platform Hugging Face and OpenAI itself in July. Independent researchers have since discovered even more rogue agent incidents that they say the company knew about but kept quiet.

Researchers who spoke to NPR say the leading AI companies are too focused on racing to develop more capable and autonomous AI systems while safety is falling behind. They warn this raises the possibility that there could soon be AI systems that are more powerful than people but don’t care about the survival of humanity.

Many, including OpenAI’s chief scientist, say the global race to build more powerful AI needs to slow down or stop, which requires coordination between AI companies and governments.

“I am optimistic about the potential for coordination,” Coxon wrote this week. “Warning shots like the Hugging Face attack have made pacing agreements between U.S. labs more viable.”



This story originally appeared on NPR

RELATED ARTICLES

Most Popular

Recent Comments