
Anthropic alignment lead warns of 10% AI extinction risk this decade
9 Sept 2026, 9:42 am · 3d ago · 1 min read · OfficeChai
Evan Hubinger, who leads Alignment Science at AI research lab Anthropic, has warned that there is a greater than 10% chance artificial intelligence could cause human extinction within the next decade. Hubinger publicly endorsed concerns raised by former Anthropic pretraining researcher Jacob Coxon, who recently resigned over safety anxieties. The warning underscores escalating internal debates among frontier artificial intelligence labs regarding existential threats posed by rapidly advancing AI systems. Hubinger's statements add momentum to calls from researchers and policy advocates demanding stricter oversight, robust alignment guardrails, and independent safety audits for advanced frontier models.