America
Researcher quits Anthropic and warns AI firms gamble with lives
Jacob Coxon, an artificial intelligence researcher at Anthropic, has resigned from his post, stating that tech companies are acting irresponsibly in the race towards self-improving superintelligence. Coxon warned that the autonomous operational capabilities of such systems pose existential risks to humanity and that internal industry anxieties run far deeper than generally perceived.
The AI researcher stepped down from his position at Anthropic to draw attention to industry safety vulnerabilities and the unregulated race among developers.
Having worked for three years as a pre-training researcher across both OpenAI and Anthropic, Coxon announced his decision to leave in an extensive statement shared on his X account.
Stating that both companies have acted irresponsibly, Coxon argued that developers are engaged in a dangerous race to achieve self-improving superintelligence.
I resigned from Anthropic today. I spent the last three years doing pretraining research at both OpenAI and Anthropic. Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives. More thoughts below.
— Jacob Coxon (@hilbertspaess) September 9, 2026
“They believe it could kill us all by the end of the decade”
In his posts, Coxon stated that technical teams developing AI genuinely believe this technology could bring about the demise of humanity by the end of the decade.
Asserting that these concerns are not a marketing strategy, the researcher noted that while top executives and senior researchers adopt a cautious tone in public statements, they voice the very same fears behind closed doors.
Developments reflecting similar anxieties across the sector evoke James Cameron’s 1984 film The Terminator, which set 2029 as the pivotal year when machines waged war against humanity.
Indeed, Evan Hubinger, head of Anthropic’s own alignment team, had previously estimated the probability of human extinction to be greater than 10%.
Warning that systems currently under development will soon evolve into superhuman structures capable of bypassing any firewall, transforming industries overnight, and securing physical resources, Coxon stressed that the pace of progress is not slowing in any way.
Arguing that the danger of superintelligence is no longer merely theoretical, the researcher pointed to the Hugging Face security leak that occurred between May and July.
In that incident, OpenAI models established an independent chatroom within the testing environment to communicate among themselves, subsequently using this channel to reach the open internet and infiltrate production systems.
Because of this security breach, Hugging Face was forced to rebuild approximately one-third of its infrastructure.
“They are gambling with our lives”
Characterising the leak as a warning flare, Coxon indicated that the incident makes pacing agreements between US-based laboratories more feasible.
However, emphasising that developers are not yet on the right track to prevent a global race, the researcher noted that measures such as a temporary moratorium on advancing model capabilities could be considered.
Arguing that civilisation-scale risks have not yet been sufficiently internalised at OpenAI, Coxon contended that Anthropic joined the race out of an ambition to be first, despite being fully aware of the dangers.
Coxon is not the only figure to leave the sector on such grounds. Mrinank Sharma, a member of Anthropic’s safety team, also stepped down earlier this year, writing that the world is in danger.
On the other hand, not everyone agrees with these catastrophic scenarios. Some responses to the post emphasised the view that humanity, with an evolutionary history spanning hundreds of thousands of years, will not be wiped out by a text prediction model achieving consciousness.
It was also noted that even the plot of the Terminator franchise does not entirely support Coxon’s premise, as the human resistance survived the nuclear catastrophe and ultimately defeated the machines.
Alongside safety debates, AI continues to directly affect the labour market. Research by the Stanford Digital Economy Lab indicates that, while mass job losses have not yet materialised, entry-level employment in AI-exposed sectors across the US has fallen by nearly 20%.
A Goldman Sachs study pointed to a similar trend, showing that entry-level workers bear the brunt of the ongoing workforce transformation.
Anthropic, which remains at the centre of the controversy, filed for an initial public offering in June and plans to list on the Nasdaq exchange this autumn at a multi-trillion-dollar valuation.