TLDRocket
Sign in

“It could kill us all”: what Anthropic’s own researchers really think about superintelligence

The New Stack Amanda Caswell Covered by 11 sources

Anthropic researchers Jacob Coxon, Evan Hubinger, and Samuel Marks warned that alignment for superintelligence is still unsolved and that current monitoring cannot be robustly verified as systems scale. Hubinger said the chance AI could kill all humans is >10% within the next decade. The disclosures point to faster capability and self-improvement loops (including Claude writing over 80% of merged code by May) that widen the gap between capability and control, increasing calls for stronger safeguards and possible slowdowns.

Why it matters

On Tuesday evening, Anthropic pretraining researcher Jacob Coxon announced on X that he’d resigned. Within hours, two of his colleagues The post “It could kill us all”: what Anthropic’s own researchers really think about superintelligence appeared first on The New Stack.

Related stories

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.