Skip to main content
3h ago

More than 1 in 10 chance AI ‘could kill all humans,’ says Anthropic safety lead after colleague quits

The departing researcher said AI companies are ‘gambling with our lives’ by racing to develop self-improving AI. The departing researcher said AI com

2 min read3 views1 tags
Originally reported bytheverge
The departing researcher said AI companies are ‘gambling with our lives’ by racing to develop self-improving AI. The departing researcher said AI companies are ‘gambling with our lives’ by racing to develop self-improving AI. A senior Anthropic safety researcher has said there is more than a 10 percent chance artificial intelligence “could kill all humans” by the end of the decade, just hours after a colleague resigned over fears the AI lab and its rivals are carelessly racing to build “superhuman systems” they cannot control. In aposton X announcing his departure, Jacob Coxon, a researcher who has trained AI systems at Anthropic, said he had quit the company over its lax approach to safety. Coxon, who previously trained systems for OpenAI, accused the two AI companies of “racing straight to self-improving superintelligence and gambling with our lives,” even though “the people building AI earnestly believe that it could kill us all by the end of the decade.” Industry insiders have long expressed concerns about the potential dangers of self-improving AI systems, which they warn could spiral out of human control in a runaway loop often described as recursive self-improvement. While not yet realized, companies are actively pursuing this goal and much of today’s AI code is written with the help of AI. In a direct response, Evan Hubinger, who leads one of Anthropic’s AI safety teams,saidhe worries about self-improving AI, adding that it “is happening faster than we thought.” He also agreed with Coxon’s characterization. “We really do earnestly believe AI could kill all humans,” he said, personally estimating the chances to be greater than one in 10 “within the next decade.” Despite this, Hubinger said Anthropic does “not yet have a plan” for ensuring advanced AI remains safe and aligned with human values and “are not clearly on track to” develop one either. Coxonsaysthe companies are “locked in a race” to develop advanced systems first so are pushing ahead “despite the risk.” Coxon’s departure marks one of the most high profile examples of an employee leaving Anthropic, a company founded by former OpenAI members following concerns over safety at the company. In recent years, multiple researchers havecited safety concernsas motivating their decision to leave OpenAI. The exchange illustrates mounting concerns within the industry about thedangers of increasingly sophisticated AI modelsand the speed at which systems are being developed, particularly as the companies prepare for anticipated IPOs. It also comes as the companies manage thefalloutfromnumerous rogue agent incidentsandhigh-profile safety warningsabout the monitorability of frontier models. A free daily digest of the news that matters most. This is the title for the native ad
#AI News
ES
Editorial StaffEditor

The Editorial Staff at AIChief is a team of professional content writers with extensive experience in AI and marketing. Founded in 2025, AIChief has quickly grown into the largest free AI resource hub in the industry.

View all posts
Reader feedback

What did you think of this story?

User Comments

Filter:
No comments yet. Be the first to comment!
Continue reading
View all news