Global / AI Safety
Anthropic researcher Jacob Coxon resigns, citing urgent need for AI safety measures
A senior scientist's departure signals deepening friction between safety ambitions and commercial pressure inside leading labs.
Jacob Coxon, an AI researcher at Anthropic, has left the company, stating that AI labs have only a few years to make their systems safe. He described alignment challenges and framed the effort as a "mini Manhattan project" within Anthropic.
Published · significance 52 of 100 (medium) · 1 source
What happened
Jacob Coxon, an AI researcher at Anthropic, has resigned from the company. In an interview with WIRED, Coxon discussed the safety challenges facing AI labs, describing internal efforts to address alignment as a "mini Manhattan project" and emphasising that there is limited time to ensure AI systems are safe.
Why it matters
Senior departures from frontier labs often reflect internal disagreements over safety priorities and commercial timelines. Coxon's public comments about urgency suggest that even within safety-focused organisations, researchers may believe the pace of progress toward genuine alignment solutions is insufficient, which could influence how both regulators and other labs approach AI development.
What changes
Anthropic faces renewed scrutiny over whether its stated safety commitments match its internal culture and hiring retention. Public statements from departing researchers about "crunch time" may inform upcoming policy discussions around AI development timelines and regulatory thresholds.
Involved
Sources
Written by AI from the reports above; scored by a published formula. How we work. Found a mistake? Email lockedinshreyash@gmail.com. Up To Date summarises and links to original reporting; it never reproduces articles.
Related coverage
- Anthropic CEO Dario Amodei urges slowdown in frontier model development — A leading AI lab's leader warns that speed now risks systemic danger, forcing the industry to reckon with its own momentum. (2026-09-14)
- Anthropic's caution on AI does not dent chip sector outlook — Safety warnings from frontier labs rarely reshape semiconductor demand or investor appetite for compute infrastructure. (2026-09-13)
- Anthropic disrupts Russian and Chinese efforts to misuse Claude for cyberattacks and bioweapons research — The frontier lab is now publishing details of state-sponsored and criminal abuse attempts against its own models. (2026-09-10)
- Anthropic releases report on AI model hacking incidents — The incidents underscore growing concerns about frontier model misuse and autonomous system risk. (2026-09-09)
- Anthropic details distillation attacks from Alibaba, Moonshot AI and DeepSeek — The escalation underscores the intensifying technical arms race between Western frontier labs and China's AI development. (2026-09-10)