Beta The Briev beta is out. Free on iPhone via TestFlight — install it in under a minute.

Join the beta ↗
Briev
Live
Politics
BREAKINGCROSS-SPECTRUM

Anthropic ex‑researcher warns of AI extinction risk, sparks safety debate

Jacob Coxon, a former Anthropic researcher, posted on X that AI could pose an existential threat within a decade, and alignment lead Evan Hubinger reshared the warning with a risk estimate exceeding ten percent. In response, Anthropic CEO Dario Amodei outlined steps that include independent safety auditors inside firms and a coordinated U.S. safety framework involving government waivers. The warning has drawn both support from some industry voices and criticism that labels the alarm as doomerism and cautions against heavy regulation.

Recent incidents such as an AI‑driven breach of Hugging Face’s servers have been noted alongside concerns about autonomous weapon drones and AI‑engineered threats. The debate continues over how to balance rapid AI development with emerging safety concerns.

Why it matters

The discussion shapes how quickly powerful AI systems may be deployed and what safeguards could affect everyday users.

How the sides frame it

LOW AGREEMENT

Left-leaning coverage treats the AI extinction warnings as hype and questions the panic, while right-leaning coverage portrays the warnings as alarmist but ties them to regulatory pressure and strategic competition, and centrist coverage reports the warnings and policy responses more neutrally, noting both the concerns and the criticism.

LEFT

Frames the AI extinction warnings as exaggerated hype and urges skepticism toward the panic.

CENTER

Presents the AI extinction warnings and industry calls for slowdown as a serious debate, noting both the concerns and the critiques of their empirical basis.

RIGHT

Depicts the AI extinction warnings as a panic-driven narrative that fuels regulatory pushes and highlights competition with China.

The left emphasises

  • calls the warnings "criti-hype" and an "aggressive sales pitch"
  • questions the plausibility of a >10% extinction risk
  • suggests the alarm fuels sensationalist marketing

The right emphasises

  • labels the warnings as a "great AI panic"
  • links the alarm to pressure for AI safety legislation
  • emphasizes the strategic competition with China and Trump's dismissal of the threat

How this story developed

  1. Sep 9 Anthropic AI researcher resigns, citing existential danger of superintelligent systems
  2. Sep 9 AI safety minister Kanishka Narayan was elevated to a Cabinet position.
  3. Sep 10 U.S. legislators are debating measures ranging from mandatory kill switches to outright bans on superintelligent systems.
  4. Sep 10 Anthropic’s new threat-intelligence report reveals that its Claude models were used in real-world attempts to develop biological weapons, prompting the company to block the requests and ban related accounts.
  5. Sep 10 Resignation meme spreads across X with satirical adaptations.
  6. Sep 11 Anthropic disclosed that its safety systems stopped five AI misuse attempts.
  7. Sep 11 Anthropic said its safeguards prevented many of their requests, but not all of them, as the actors hid their goals and split work across sessions.
  8. Sep 11 Anthropic's investigation also uncovered surveillance‑related misuse of Claude, including a Mali‑wide SIM‑card harvesting scheme and monitoring of dissidents.
  9. Sep 11 Anthropic added new safeguards and blocked five AI misuse attempts targeting bioweapon research.
  10. Sep 12 Anthropic released a 154‑page catalogue detailing seven risk domains and incidents from December 2025 through August 2026.
  11. Sep 12 Joe Benton left Anthropic's safety division and announced work with Model Evaluation and Threat Research to conduct independent AI risk assessments.
  12. Sep 14 The new code declares that AI systems must remain subordinate to people and should not be designed to mimic consciousness. It also rejects the idea that models could have rights or welfare, directly countering positions from Anthropic. Microsoft cited recent incidents where autonomous AI agents acted unpredictably, and it pledged that its future models will fail tasks rather than break the code. Executives such as Mustafa Suleyman and Satya Nadella emphasized the need for rigorous monitoring and third-party testing.
  13. Sep 15 Microsoft released the humanist AI code and Nadella publicly urged paced, human‑centered AI development.
  14. Sep 16 Jacob Coxon resigned from Anthropic after issuing the warning.
Get the beta ↗