“It could kill us all”: what Anthropic’s own researchers really think about superintelligence
The New Stack Amanda Caswell ● Covered by 11 sources
Anthropic researchers Jacob Coxon, Evan Hubinger, and Samuel Marks warned that alignment for superintelligence is still unsolved and that current monitoring cannot be robustly verified as systems scale. Hubinger said the chance AI could kill all humans is >10% within the next decade. The disclosures point to faster capability and self-improvement loops (including Claude writing over 80% of merged code by May) that widen the gap between capability and control, increasing calls for stronger safeguards and possible slowdowns.
Why it matters
On Tuesday evening, Anthropic pretraining researcher Jacob Coxon announced on X that he’d resigned. Within hours, two of his colleagues The post “It could kill us all”: what Anthropic’s own researchers really think about superintelligence appeared first on The New Stack.
Related stories
Anthropic researcher quits with a warning: Self-improving AI could "kill us all"
Ars Technica · 16 hours ago ·
8
More than 1 in 10 chance AI ‘could kill all humans,’ says Anthropic safety lead after colleague quits
The Verge · 1 day ago ·
19
Superintelligence is coming. Should we let it?
TechCrunch · 17 hours ago ·
31