Anthropic researchers warn AI could threaten humanity
Warnings from inside two of the world's leading artificial intelligence companies have intensified a political debate in Britain over the risks of advanced AI, after a researcher resigned from Anthropic and a colleague said he believes the technology could pose a species-ending threat within a decade. Jacob Coxon, a researcher who previously worked at both Anthropic and OpenAI, announced his resignation on Wednesday in a series of posts on X that drew millions of views. He wrote that neither company was acting responsibly and that both were racing toward self-improving superintelligence while gambling with people's lives. The people building AI earnestly believe it could kill everyone by the end of the decade, he said, adding that this was not a marketing stunt. Evan Hubinger, Anthropic's alignment science lead, responded on X that he agreed with Coxon. He said he personally believes there is a greater than 10 percent chance AI could kill all humans within the next decade, and that while Anthropic is trying its best, the company does not yet have a plan to solve alignment for superintelligence and is not clearly on track to do so.