← All stories
● Covered by 3 sources · 3 reportsHigh impact3 negative

Anthropic Researcher Warns Over 10% Chance AI Could 'Kill All Humans' Within a Decade

🔄 Updated 58m ago
New to BrevFeed? We gather this story from every outlet covering it into one summary — ranked by real-world impact, not just the latest headline — so you never miss what matters. What is BrevFeed? →

Key points

  • Anthropic researcher Evan Hubinger stated over 10% chance AI could 'kill all humans'.
  • Hubinger's warning refers to risks within the next decade.
  • Jacob Coxon resigned from Anthropic, citing irresponsible AI development.
  • Coxon previously worked at OpenAI and Anthropic.
  • Concerns focus on self-improving superintelligence and existential risk.

AI Existential Risk Warning

Evan Hubinger, a safety researcher at Anthropic, has publicly stated that there is a greater than 10% chance artificial intelligence could lead to human extinction within the next ten years. Hubinger noted that while the risk from current models is low, he is concerned about the technology's potential to rapidly improve itself to an existential threat level.

Researcher Resignation and Accusations

This warning follows the resignation of Jacob Coxon, another researcher from Anthropic. Coxon, who previously worked at both OpenAI and Anthropic, accused both companies of acting irresponsibly. He stated that they are "racing straight to self-improving superintelligence and gambling with our lives."

Coxon emphasized the power of this technology, warning that AI systems could soon become superhuman, capable of hacking anything, revolutionizing fields overnight, and acquiring real power and resources.

Internal Concerns and Company Stance

These comments underscore increasing concerns among those directly involved in AI development regarding the technology's potential to get out of control. Hubinger stated that he believes Anthropic is "trying its best" but acknowledged that there is currently no clear plan to solve alignment for superintelligence.

The BBC reported that Anthropic withheld its latest model from the AI Safety Institute, which conducts AI testing. Anthropic has been approached for comment regarding these developments.

The Concept of Self-Improving AI

Self-improvement, or recursive self-improvement, refers to the idea that AI systems can enhance themselves with minimal human intervention. While this capability is not yet realized, AI laboratories are actively working towards achieving this goal. The potential for such systems to rapidly advance without human oversight is a central point of concern for researchers like Hubinger and Coxon.

✨ This summary was generated by AI from the outlets' reporting listed below. It is not independently verified and may contain errors — check the original sources. How BrevFeed works →

The daily brief

One email each morning: the day's tech stories, clustered across outlets and summarized. No account needed.

One email a day. Unsubscribe in one click, any time.

Today's brief

Spend a few minutes, get the whole day. Every topic's top stories in one hands-free rundown — listen, watch, or read the transcript.

~8 min · 6 stories · Sep 08

▶ Play today's brief Listen on Spotify

New every morning, and the back catalogue is archived by date.

How outlets covered it

An Anthropic safety researcher, Evan Hubinger, stated there is over a 10% chance AI could lead to human extinction within the next decade, confirming concerns raised by former Anthropic and OpenAI researcher Jacob Coxon. Coxon resigned, accusing both companies of irresponsibly pursuing self-improving superintelligence without adequate safety measures. This highlights growing internal concerns within leading AI development firms regarding the potential risks of advanced AI.

Evan Hubinger, a safety researcher at Anthropic, stated there is a greater than 10% chance AI could lead to human extinction within the next decade, citing concerns about AI's rapid self-improvement capabilities. This warning follows reports that Anthropic withheld its latest model from the AI Safety Institute and aligns with increasing alarms from AI leaders about controlling advanced AI systems.

An Anthropic safety researcher stated there is over a 10% chance AI could "kill all humans" within the next decade, following a colleague's resignation over concerns about AI labs "gambling with our lives." This highlights growing internal concerns within AI development companies regarding the potential for AI to become uncontrollable and pose an existential threat.