Anthropic researcher quits, warns against self-improving AI's existential risks
The resignation follows growing concerns about AI's potential to cause human extinction. The researcher cited risks tied to self-improving models and called for a pause in development.

Jacob Coxon, a former researcher at Anthropic, has resigned and issued a stark warning about the dangers of self-improving artificial intelligence. In a public statement, he accused Anthropic and other leading AI firms of prioritizing rapid development over safety, risking catastrophic outcomes for humanity. Coxon's departure has sparked renewed debate about the ethical and existential implications of advancing AI technologies without adequate safeguards.
Coxon spent three years working on pretraining research at both OpenAI and Anthropic before resigning. He criticized the companies for their approach to AI development, arguing that they are racing toward self-improving superintelligence without fully understanding the risks. His concerns align with those of other industry insiders and policymakers who have called for a slowdown in AI progress to address potential dangers.
Evan Hubinger, who leads Anthropic’s Alignment Science team, has echoed Coxon’s warnings, estimating that AI could pose a greater than 10% chance of causing human extinction within the next decade. This estimate, based on ongoing research and analysis, has intensified pressure on AI firms to adopt more cautious strategies. Hubinger’s comments highlight the growing consensus among experts that the risks of uncontrolled AI development are too great to ignore.
The resignation and associated warnings have raised questions about the long-term costs of unchecked AI innovation. Industry stakeholders are now grappling with the potential for increased regulatory scrutiny, higher development costs, and the risk of vendor lock-in as companies seek to balance progress with safety. Market reactions have been mixed, with some investors expressing concern over the implications for AI’s future trajectory.
As the debate over AI safety continues, the broader implications for the industry remain unclear. The resignation underscores the urgency of addressing alignment challenges and ensuring that AI development is guided by ethical considerations. Without a coordinated effort to manage risks, the potential for unintended consequences could escalate, with significant consequences for both the technology sector and society at large.
Sources
- https://arstechnica.com/ai/2026/09/anthropic-researcher-quits-with-a-warning-self-improving-ai-could-kill-us-all/
- https://techcrunch.com/2026/09/09/gambling-with-our-lives-anthropic-researcher-quits-warns-against-self-improving-ai/
- https://www.cbsnews.com/news/ai-kill-humans-anthropic-researcher-more-than-ten-percent-chance/
- https://www.makeuseof.com/an-anthropic-researcher-thinks-ai-could-kill-us-all-in-a-few-years/