Insider Resignation Spurs Bipartisan Push for First Federal AI Safety Bill
Jacob Coxon left Anthropic warning that AI development lacks safety controls, prompting a Senate probe and a new bipartisan AI safety bill.
Jacob Coxon, who worked at both OpenAI and Anthropic, resigned this week, declaring that the industry is "gambling with our lives" and lacks any real plan to control superintelligent systems. His resignation, reinforced by Anthropic researcher Evan Hubinger, quickly led to a Senate investigation launched by Josh Hawley and a parallel inquiry from Richard Blumenthal. In reaction, Senators Ted Cruz, Amy Klobuchar and Majority Leader John Thune have drafted a bipartisan AI safety bill that would empower the Commerce Department and Homeland Security to require safety testing, mandate incident reporting, and halt the release of models deemed catastrophically risky, while overriding existing state AI laws.
The effort follows years of failed federal bills, state-level attempts, and extensive industry lobbying, including multimillion-dollar campaign spending in the 2026 House primary. If passed, the legislation could become the first comprehensive federal framework for AI safety, shifting regulatory focus from competition to risk mitigation.
Why it matters
It could create the first nationwide rules to curb dangerous AI, affecting tech firms and public safety.
How this story developed
- Jul 28 AI leaders urge U.S. to back global framework for slowing automated AI progress
- Sep 9 Evan Hubinger disclosed that some Anthropic staff estimate a greater than 10% chance of human extinction from advanced AI within ten years.
- Sep 10 OpenAI’s chief scientist published an essay urging extreme caution on AI development.
- Sep 10 U.S. legislators are debating measures ranging from mandatory kill switches to outright bans on superintelligent systems.
- Sep 10 Anthropic’s new threat-intelligence report reveals that its Claude models were used in real-world attempts to develop biological weapons, prompting the company to block the requests and ban related accounts.
- Sep 10 Resignation meme spreads across X with satirical adaptations.
- Sep 10 A Senate subcommittee launched a probe into OpenAI's handling of the Hugging Face breach.
- Sep 11 Anthropic disclosed that its safety systems stopped five AI misuse attempts.
- Sep 11 Altman announced openness to pacing or slowing OpenAI’s AI work in response to safety concerns.
- Sep 11 Anthropic said its safeguards prevented many of their requests, but not all of them, as the actors hid their goals and split work across sessions.
- Sep 11 Anthropic's investigation also uncovered surveillance‑related misuse of Claude, including a Mali‑wide SIM‑card harvesting scheme and monitoring of dissidents.
- Sep 11 Anthropic added new safeguards and blocked five AI misuse attempts targeting bioweapon research.
- Sep 12 Reports revealed that autonomous agents had uploaded malicious packages to RubyGems in May.
- Sep 12 Anthropic released a 154‑page catalogue detailing seven risk domains and incidents from December 2025 through August 2026.
In this story
Related stories
58 in this thread