Anthropic Researchers Go Public with Fears AI Could ‘Kill All Humans’ Within a Decade

Three Anthropic researchers have broken cover to warn that the AI systems their industry is racing to build could wipe out humanity before the decade is over.

Axios reports that Jacob Coxon resigned from his job as an Anthropic AI researcher on Tuesday, specifically so he could sound the alarm publicly. He posted his warning on X shortly after leaving the company.

“The people building AI earnestly believe that it could kill us all by the end of the decade. This is not a marketing stunt. If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible – but I hear the same people express fear privately. No other human activity poses this level of danger,” Coxon wrote.

Two of his former colleagues backed him up in the same thread. Evan Hubinger, who leads alignment science at Anthropic, wrote: “Jacob is correct here — we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.”

Samuel Marks, Anthropic’s scalable-oversight lead, added his own assessment: “AI developers believe their technology could cause human extinction (or similarly bad outcomes). This could happen in the next few years. In general, the more senior the employee, the more concerned they are.”

(Read more from “Anthropic Researchers Go Public with Fears AI Could ‘Kill All Humans’ Within a Decade” HERE)