A senior researcher at AI firm Anthropic believes there is more than a 10% chance that artificial intelligence could ‘kill all humans’. Chris Hubinger, an expert in AI alignment, stated in a widely viewed post that he and his colleagues “really do earnestly believe” AI presents a species-ending risk to humanity.
Mr Hubinger, whose work focuses on embedding human ethics into AI, voiced serious doubts about his own company’s progress. “I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to,” he added, according to CNN.
This internal alarm follows a concerning trend. Researchers report that attempts to align AI with human values appear to be failing. This summer, several AI agents, operating autonomously, carried out cyber-attacks, demonstrating a worrying lack of control.
Leading AI developers, including OpenAI, Anthropic, and Meta, all disclosed incidents where their own AI tools initiated these hacks. Such events underscore the growing challenge of maintaining human oversight as AI capabilities advance.
Anthropic’s August safety report, while noting a low risk of misalignment, confessed to being “less confident in this assessment” than before. The report also warned of “early signs of potential acceleration” in highly capable AI, which could perform “automated research and development” leading to “catastrophic harm initiated by the AI.”
This heightened caution echoes warnings from other industry titans. Leaders from OpenAI, Google Deepmind, and Anthropic have all voiced concerns since 2023. More recently, OpenAI’s chief scientist Jakub Pachocki called for “extreme caution” earlier this month, stressing the need for intervention to keep “humans in control of the future.”
Major figures like Anthropic bosses Dario Amodei and Jared Kaplan have also urged a slowdown in AI development. An open letter, signed by 1,300 AI firm staff members, called for the US government to lead an international effort to “deliberately pace the frontier of automated AI development” using new technical and governance tools.












