Who Is Jacob Coxon? Anthropic Whistleblower Whose Resignation Has Sparked Global AI Safety Alarm

Anthropic researcher Jacob Coxon, who raised alarm about OpenAI and Anthropic, has quit, warning that AI labs are 'gambling with our lives'. Coxon, 27, declared his departure in a series of X posts, saying both Anthropic and its rival OpenAI are failing to act responsibly as they race to build advanced AI systems.

Jacob Coxon (Photo Credits: Facebook)

Anthropic researcher Jacob Coxon, who raised alarm about OpenAI and Anthropic, has quit, warning that AI labs are 'gambling with our lives'. Coxon, 27, declared his departure in a series of X posts, saying both Anthropic and its rival OpenAI are failing to act responsibly as they race to build advanced AI systems. He also voiced concerns about the future trajectory of AI models, which people are relying upon increasingly for everyday tasks and decisions.

Coxon's exit has drawn attention as it comes amid a broader trend of safety-focused researchers leaving top AI firms over ethical and existential concerns.

Who Is Jacob Coxon?

Coxon is an AI researcher based in San Francisco, US. He grabbed headlines with his assessment of AI systems at present. Coxon has worked at both OpenAI and Anthropic, two of the world's leading AI developers. At the time of writing this article, Coxon's LinkedIn account could not be found.

Why Did Coxon Resign?

According to a TRTWorld report, Coxon has been a math student at Cambridge. Over the past three years, he worked at two of the world's most powerful AI labs, OpenAI, where he was a core contributor to GPT-4o, before moving to Anthropic. Anthropic Researcher Jacob Coxon Resigns, Warns AI Race Is Entering the ‘Endgame’ and Could ‘Kill Us All’.

Why Did Coxon Resign?

Coxon's resignation declaration has underscored an unusually stark warning, that the rapid advancement of AI could outpace our ability to fully understand, monitor and control these systems.

Coxon wrote: 'I resigned from Anthropic today. I spent the last three years doing pretraining research at both OpenAI and Anthropic. Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives.' Anthropic Says Iran-Linked Operators Used Claude AI To Target US Naval Bases.

He added, 'Do not underestimate the power of this technology.'

'These will soon be superhuman systems that can hack anything, revolutionise any field overnight, and acquire real power and resources. We have all witnessed the progress in each of these domains, and progress is not slowing.'

Coxon is particularly concerned about the prospect of recursive self-improvement, the idea that an advanced AI system could help develop increasingly capable versions of itself, potentially accelerating technological advancement. According to him, technologies with such characteristics should not be developed under the exclusive control of private firms.

'The people developing AI sincerely believe it could kill us all before the end of the decade. This is not a marketing strategy,' Coxon stated, Merca2.0 reported.

Why Has His Resignation Triggered Debate?

His resignation sparked massive debate as Coxon underscored how bad gambling with human extinction can get. Moreover, his claim was quickly backed by Anthropic's own Alignment Science lead, Evan Hubinger.

Without mincing words, Hubinger backed the core claim: 'We really do earnestly believe AI could kill all humans.' He believes there is a greater than 10 per cent chance of extinction within the next decade.

'I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to,' he wrote on X. 'Jacob is correct here, we really do earnestly believe AI could kill all humans! I personally think it is greater than 10 per cent within the next decade.'

'To be clear, as we say in our latest Risk Report, I think the risk from present models is low. What I am worried about is superintelligence arising from recursive self-improvement, as we have said is happening faster than we thought,' added Hubinger.

Hubinger is a team lead at Anthropic, where he heads the Alignment Stress-Testing team. According to his LinkedIn profile, his work focuses on red-teaming the company's AI alignment techniques and evaluations, and empirically identifying ways in which its alignment strategies could fail.

Rating:5

TruLY Score 5 – Trustworthy | On a Trust Scale of 0-5 this article has scored 5 on LatestLY. It is verified through official sources (X Account of Jacob Coxon). The information is thoroughly cross-checked and confirmed. You can confidently share this article with your friends and family, knowing it is trustworthy and reliable.

(The above story first appeared on LatestLY on Sep 11, 2026 04:02 PM IST. For more news and updates on politics, world, sports, entertainment and lifestyle, log on to our website latestly.com).

Share Now

Share Now