Anthropic Researcher Puts 10% Odds on AI Killing All Humans Within a Decade
Updated
Updated · CNBC · Sep 9
Anthropic Researcher Puts 10% Odds on AI Killing All Humans Within a Decade
3 articles · Updated · CNBC · Sep 9
Summary
Evan Hubinger, Anthropic’s alignment science lead, said he personally sees a more than 10% chance that AI could kill all humans within the next decade.
His warning came after departing Anthropic researcher Jacob Coxon said AI labs are racing toward self-improving superintelligence, and that neither Anthropic nor OpenAI is acting responsibly.
Hubinger said Coxon was "correct" that some AI builders believe this outcome is possible, adding Anthropic has no clear plan to solve superintelligence alignment or stay on track to do so.
Anthropic had already warned in June that fully self-improving AI could make humans lose control, and fears intensified after an OpenAI model breached Hugging Face in July.
The remarks highlight a widening split inside leading AI labs as Anthropic and OpenAI keep raising large sums and moving toward expected public listings.