Departing Anthropic Researcher Says the AI Race Has No Safe Winner

Jacob Coxon spent three years doing pretraining work, initially for OpenAI before switching to Anthropic. This week, he quit the industry altogether.

Coxon, 27, made that switch earlier this year, drawn in part by Anthropic’s reputation as the more safety-conscious of the two labs. He still thinks the safety effort there is real — he just no longer believes any one company can hold that line while locked in competition with everyone else.

He laid out his reasoning in a lengthy post on X. Neither lab is taking the risk seriously enough, he wrote, and both are “gambling with our lives” by pushing toward machines that can improve themselves faster than anyone can properly supervise. People inside these labs have started using shorthand for how close this feels — terms like “crunchtime” and “endgame” come up often.

Coxon says colleagues who choose careful, reassuring words for reporters describe something very different once the cameras are off — a real, not hypothetical, chance that AI could end human civilization within the decade. Nothing else people build, in his view, comes anywhere close to that kind of stakes.

Altamira Gold Corp. — sponsored Sponsored · Altamira Gold Corp.

OpenAI, by his account, simply hasn’t grasped the stakes as an organization. Anthropic is different — its people understand the risk, but the company has talked itself into believing it can’t afford to slow down, on the theory that pulling back only clears the way for a competitor with fewer scruples. He calls the decision to keep going a hubristic gamble that shouldn’t be made unilaterally inside a single company.

Engineers are building this technology on ordinary company laptops in San Francisco, he argued, not with anything like the physical security and secrecy of the original Manhattan Project — a comparison meant to underline how casually he thinks the industry treats the stakes.

Coxon pointed to a July security incident as evidence the risk is already showing up. OpenAI later confirmed that a mix of its own models, running with their usual safety limits dialed down to test hacking ability, strung together a zero-day flaw with stolen login credentials to reach Hugging Face‘s production systems and pull information the evaluation was supposed to keep out of reach. Hugging Face’s security staff caught the intrusion and shut it down.

Not the first warning

Coxon isn’t the first Anthropic researcher to leave publicly alarmed. In February, Mrinank Sharma, who led the company’s safeguards research team, resigned with a letter saying “the world is in peril.” Coxon’s departure also comes two days after OpenAI Chief Scientist Jakub Pachocki published an essay called “An Alien Mind,” arguing that no lab, including his own, has cracked the alignment and monitoring problem well enough to justify racing ahead at today’s pace.

Read: Anthropic Safety Head Abandons Tech for Poetry, Warns of Global Crises

Evan Hubinger, who leads Anthropic’s alignment science work, replied to Coxon on X within about ninety minutes. He didn’t dispute the claim about private beliefs. He confirmed it. He and his colleagues “really do earnestly believe AI could kill all humans,” he wrote, putting his own odds above 10% over the next decade and admitting Anthropic still has no real plan for keeping a superintelligent system aligned.

Hubinger had every reason to downplay Coxon’s warning and instead handed it more weight, putting a specific number on a claim most executives would rather leave vague.

The IPO math

Anthropic raised $65 billion in May at a $965 billion valuation, on run-rate revenue that had just crossed $47 billion, months after a $30 billion raise in February valued it at $380 billion. The company quietly submitted IPO paperwork to regulators around June 1, and a Wall Street debut could come before winter, with Goldman Sachs (NYSE: GS), JPMorgan (NYSE: JPM) and Morgan Stanley (NYSE: MS) all said to be in the running to lead it. No Anthropic executive has confirmed a target valuation, though investors have floated figures near $2 trillion.

Read: Anthropic Files Confidential S-1 as AI IPO Wave Nears $3 Trillion 

Information for this briefing was found via the sources and the companies mentioned. The author has no securities or affiliations related to this organization. Not a recommendation to buy or sell. Always do additional research and consult a professional before purchasing a security. The author holds no licenses.

Leave a Reply

Video Articles

A $2.2B Gold Project Is Outgrowing Its Plan | Michael Henrichsen – Gold X2 Mining

Pay for the Copper, Get the Gold Free | Rob McEwen – McEwen Inc

This Gold Discovery Was Already Huge. Now It’s Becoming a Monster. | Goliath Resources

Recommended

Golden Cariboo’s First Quesnelle Resource Estimate Tallies 1.19 Million Gold Equivalent Ounces

Brixton Wraps Camp Creek Drilling With 17.58 Metres of 1.47 g/t Gold Equivalent

Related News

Anthropic Suffers Two Data Exposures in Five Days, Revealing Unreleased Model and Full Source Code

Anthropic leaked 512,000 lines of proprietary source code and details of an unannounced AI model...

Wednesday, April 1, 2026, 07:35:27 AM

Anthropic Walks Away From $6 Billion Decart Deal After Due Diligence

Anthropic has abandoned a potential $6 billion acquisition of Decart AI after conducting due diligence,...

Tuesday, September 8, 2026, 10:29:00 AM

Authors Sue Anthropic for Alleged ‘Large-Scale Theft’ of Copyrighted Books

AI startup Anthropic is facing a class-action lawsuit alleging copyright infringement. Filed on Monday in...

Wednesday, August 21, 2024, 04:14:00 PM

The Pentagon Banned Anthropic — Then Gave OpenAI the Same Deal

President Donald Trump ordered every federal agency to cut ties with Anthropic on Friday, and...

Monday, March 2, 2026, 11:07:00 AM

Anthropic Safety Head Abandons Tech for Poetry, Warns of Global Crises

Mrinank Sharma, who led AI safety research at Anthropic, resigned Monday, warning that “the world...

Wednesday, February 11, 2026, 12:54:00 PM