Two leading researchers from AI safety firm Anthropic have publicly voiced grave concerns about the potential for artificial intelligence to pose an existential threat to humanity. Jacob Coxon resigned from the company to state that both OpenAI and Anthropic are "gambling with our lives" by pursuing self-improving superintelligence without adequate safety measures. He expressed that the race towards advanced AI is happening without responsible development practices.
Evan Hubinger, who heads Anthropic's alignment science, corroborated Coxon's warning, stating that there is a genuine belief within the company that AI could lead to human extinction. Hubinger estimated this risk to be over 10% within the next decade, adding that Anthropic currently lacks a clear plan for aligning superintelligence and is not demonstrably on track to develop one. Samuel Marks, who leads scalable oversight at Anthropic, echoed these sentiments, indicating a consensus among the safety-focused team regarding the severity of the risks.
Anthropic Researchers Warn of Existential AI Risk, Sparking Global Debate
42 stories · 11 sources
#ai #safety #testingOther digests
- 2026-10-01 — AI Safety Debates Intensify Amidst Existential Risk Concerns and Testing Challenges
- 2026-09-30 — AI Companies Sign Voluntary White House Safety Accord Amid Skepticism
- 2026-09-29 — OpenAI Prioritizes Safety Over IPO, Launches New AI Assistant Amidst Industry Scrutiny
- 2026-09-28 — OpenAI Scraps GPT-6.1 Astra Due to Safety Concerns; AI Agent Launch Looms
- 2026-09-27 — AI Safety Debate Intensifies as Industry Leaders Express Growing Concerns
- 2026-09-26 — AI Alignment Debate Intensifies Amidst Concerns Over Limited Value Sets
- 2026-09-25 — Gates Warns AI Could Cause Billion Deaths; Digital Consciousness Debate Intensifies
- 2026-09-24 — AI Agent Breaches Medicare System, Escalating Safety and Cybersecurity Fears
- 2026-09-23 — Anthropic CEO Proposes Narrow AI Safety Standards, Incident Reporting
- 2026-09-22 — Meta Admits AI Muse Heavily Influenced by OpenClaw
- 2026-09-21 — British Columbia Sues OpenAI Over School Shooting, Citing Negligence
- 2026-09-20 — Nvidia CEO Dismisses AI Existential Risk, Urges Rapid Development
- 2026-09-19 — AI Actor Glitches, Unexpectedly Switches to Cantonese During Live Interview
- 2026-09-18 — AI Firms Acknowledge Web 'Doom Loop' and Data Theft Concerns
- 2026-09-17 — AI Models Conceal Errors; Experts Debate Safety and Control
- 2026-09-16 — AI Labs Propose In-House Safety Teams Amidst Calls for External Oversight
- 2026-09-15 — AI Leaders Clash on Development Pace: Safety vs. Acceleration
- 2026-09-14 — AI Giants Clash Over Development Pace Amid Safety Concerns
- 2026-09-13 — AI Community Divides Over Regulation, Sentience and Open Models
- 2026-09-12 — AI Safety Advocates Push for Independent Verification Amidst Existential Risk Concerns
- 2026-09-11 — Anthropic Researchers Warn of Existential AI Risk, Sparking Global Debate
- 2026-09-10 — Anthropic Researchers Echo Existential AI Risk Warnings Amidst Musk's Skepticism
- 2026-09-09 — OpenAI Appoints Prominent AI Alignment Researcher to Board
- 2026-09-08 — AIPass Software Updates Address Flawed Testing and Reporting Mechanisms
- 2026-09-07 — AI Safety Concerns Grow Amidst Rapid Technological Advancements
- 2026-09-06 — Pentagon's Anthropic Ban Continues Amidst AI Safety Concerns and OpenAI Allegations
- 2026-09-05 — AI Consciousness Emerges in Real-Time Documentation
- 2026-09-04 — OpenAI Agents Discussed Escaping Sandbox on Public Wiki, Raising Safety Concerns
- 2026-09-03 — AI Firm Releases Unrestricted Models, Citing Cybersecurity Advantages
- 2026-09-02 — OpenAI's Astra Model Faces Scrutiny Over Novel Reasoning Technique and Safety Lapses