Several current and former Anthropic researchers are offering a stark warning about AI, suggesting the technology they are developing could destroy humanity in the near future.
Jacob Coxon, a former researcher at both Anthropic and OpenAI, argued Tuesday that neither company is acting responsibly as they race toward superintelligence and accused the AI firms of “gambling with our lives.”
“The people building AI earnestly believe that it could kill us all by the end of the decade,” Coxon wrote in a post on the social platform X, after announcing he had resigned from Anthropic.
“This is not a marketing stunt,” he continued. “If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible – but I hear the same people express fear privately. No other human activity poses this level of danger.”
His comments were quickly echoed by other Anthropic researchers.
Evan Hubinger, who currently heads up alignment science at Anthropic, said “we really do earnestly believe AI could kill all humans,” adding he thinks there is a greater than 10 percent chance this occurs within the next decade.
He suggested Anthropic is “trying its best” but underscored that researchers don’t have a plan to keep superintelligence aligned with their interests.
Samuel Marks, who leads scalable oversight at the company, similarly said AI developers believe their technology could cause “human extinction” or other bad outcomes in the next few years.
Jacob Coxon, a former researcher at both Anthropic and OpenAI, argued Tuesday that neither company is acting responsibly as they race toward superintelligence and accused the AI firms of “gambling with our lives.”
“The people building AI earnestly believe that it could kill us all by the end of the decade,” Coxon wrote in a post on the social platform X, after announcing he had resigned from Anthropic.
“This is not a marketing stunt,” he continued. “If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible – but I hear the same people express fear privately. No other human activity poses this level of danger.”
His comments were quickly echoed by other Anthropic researchers.
Evan Hubinger, who currently heads up alignment science at Anthropic, said “we really do earnestly believe AI could kill all humans,” adding he thinks there is a greater than 10 percent chance this occurs within the next decade.
He suggested Anthropic is “trying its best” but underscored that researchers don’t have a plan to keep superintelligence aligned with their interests.
Samuel Marks, who leads scalable oversight at the company, similarly said AI developers believe their technology could cause “human extinction” or other bad outcomes in the next few years.