Anthropic Researcher Resigns Over AI Safety Concerns, Warns of Existential Risks
Jacob Coxon, a 27-year-old British researcher at Anthropic, has publicly resigned, expressing grave concerns about the rapid development of artificial intelligence and its potential to lead to human extinction. His departure, which occurred around September 8-9, highlights a growing unease among AI professionals regarding the manageability of advanced AI systems.
Coxon’s resignation is particularly alarming as it has garnered support from his colleagues at Anthropic. Evan Hubinger, the Alignment Science Lead, echoed Coxon’s fears on social media, estimating a greater than 10% chance of AI causing significant human fatalities within the next decade. He also noted that Anthropic currently lacks a clear strategy for aligning with superintelligent systems. Samuel Marks, another researcher, shared similar concerns, indicating that many senior employees at the company are worried about potential extinction-level risks.
Coxon criticized the language used in AI development, such as “crunchtime” and “endgame,” arguing that it frames the pursuit of superintelligence as an urgent competition rather than a significant risk that needs careful management. This resignation is not an isolated incident; it follows the departure of Mrinank Sharma, who led Anthropic’s Safeguards Research team, earlier in the year due to similar concerns about escalating dangers posed by AI.
The irony of the situation is notable, as Anthropic was founded by former OpenAI staff who aimed to prioritize safety in AI development. Now, their own safety researchers are leaving due to perceived negligence in addressing these critical issues. Coxon specifically pointed to international competition, particularly from Chinese AI labs, as a factor exacerbating the urgency to advance AI technologies without adequate safety measures.
FAQ
Why did Jacob Coxon resign from Anthropic?
Jacob Coxon resigned due to grave concerns about the rapid development of artificial intelligence and its potential to lead to human extinction, highlighting a growing unease among AI professionals regarding the manageability of advanced AI systems.
What are the concerns raised by Evan Hubinger regarding AI?
Evan Hubinger, the Alignment Science Lead at Anthropic, expressed concerns about the potential for AI to cause significant human fatalities, estimating a greater than 10% chance of this occurring within the next decade, and noted a lack of a clear strategy for aligning with superintelligent systems.
How did Coxon criticize the language used in AI development?
Coxon criticized terms like 'crunchtime' and 'endgame,' arguing that they frame the pursuit of superintelligence as an urgent competition rather than highlighting the significant risks that require careful management.
What prior incident preceded Coxon's resignation at Anthropic?
Coxon's resignation followed the departure of Mrinank Sharma, who led Anthropic’s Safeguards Research team, earlier in the year due to similar concerns about the escalating dangers posed by AI.
What irony is noted regarding Anthropic's founding and current situation?
The irony lies in the fact that Anthropic was founded by former OpenAI staff who aimed to prioritize safety in AI development, yet now their own safety researchers are resigning due to perceived negligence in addressing critical safety issues.
Comments
Comments are moderated before publish.
No comments yet — be the first.