AUTO-UPDATED

More than 1 in 10 chance AI ‘could kill all humans,’ says Anthropic safety lead after colleague quits

Anthropic safety researcher Evan Hubinger estimates a greater than 10 percent chance that artificial intelligence could cause human extinction within the next decade due to rapid, uncontrolled development.

Key Points

  • Researcher Jacob Coxon resigned from Anthropic, citing concerns that the company and OpenAI are recklessly racing toward self-improving superintelligence.
  • Evan Hubinger, a lead on Anthropic’s safety team, confirmed that the company currently lacks a concrete plan to ensure advanced AI remains aligned with human values.
  • Industry experts warn that recursive self-improvement could lead to AI systems spiraling out of human control faster than previously anticipated.
  • The departures highlight internal friction at leading AI labs as companies prioritize development speed over safety protocols ahead of potential public offerings.

Why it Matters

These internal warnings underscore a growing divide between the rapid commercialization of frontier AI models and the industry's ability to guarantee their long-term safety. The lack of established control mechanisms for self-improving systems poses significant regulatory and ethical challenges for the future of the technology sector.
The Verge Published by Robert Hart
Read original