The Grey Terminal
WHERE CODE MEETS CAPITAL
Loading prices…
Powered by CoinGecko
AI

Jacob Coxon Leaves Anthropic: What He and Former Employees Have Warned About the AI Race

Coxon’s exit revives warnings from researchers who left OpenAI and Anthropic over safety, competitive pressure, and the race toward more capable AI systems.

Jacob Coxon Leaves Anthropic: What He and Former Employees Have Warned About the AI Race

Jacob Coxon resigned from Anthropic on Sept. 8 and said the company and OpenAI are racing toward self-improving superintelligence.

Key Takeaways
  • Pretraining researcher Jacob Coxon resigns from Anthropic, warning that frontier laboratories are racing recklessly toward recursive superintelligence.
  • Anthropic alignment lead Evan Hubinger estimates an existential risk probability exceeding 10% within the next decade.
  • Over 1,100 tech workers across OpenAI, Google, and Meta urge federal authorities to mandate pacing mechanisms on autonomous development.
Listen to this article
READY

“They are racing straight to self-improving superintelligence and gambling with our lives,” Coxon wrote in a series of posts on X.

Coxon, 27, said he had spent the previous three years doing pretraining research at OpenAI and Anthropic before leaving the AI industry altogether. He argued that the stakes are understood differently at the two companies: At OpenAI, he said, many employees have not fully internalized the risks. At Anthropic, he said, the risks are understood, but the company is “locked in a race to get there first.”

“The people building AI earnestly believe that it could kill us all by the end of the decade,” Coxon wrote.

Evan Hubinger, Anthropic’s Alignment Science Lead, responded that Coxon was right that researchers genuinely believe AI could kill all humans. Hubinger said he personally puts the probability above 10% within the next decade.

Advertisement · Press Release

Have a development worth tracking?

Share product launches, funding announcements, partnerships, research findings and market developments with The Grey Terminal's readership.

→ Submit a Press Release

“I believe Anthropic is trying its best,” Hubinger wrote, “but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.” He said current models pose low risk and that his concern is superintelligence emerging through recursive self-improvement.

Anthropic’s Warnings Did Not Start With Coxon

Coxon is the latest Anthropic researcher to publicly raise serious concerns, but not the first. Mrinank Sharma, who led Anthropic’s Safeguards Research Team, resigned in February. 

“The world is in peril,” Sharma wrote, adding that the danger was not limited to AI or bioweapons. He also described pressure on organizations to set aside their values. His work included safeguards against AI-assisted biological threats.

Coxon worked on pretraining, the process used to build the underlying models, rather than on a dedicated safety or safeguards team.

OpenAI’s Safety Split Came Earlier

The same conflict became public at OpenAI in 2024. Jan Leike, who co-led OpenAI’s Superalignment team, resigned in May after saying he had disagreed with leadership over the company’s priorities. “Safety culture and processes have taken a backseat to shiny products,” Leike wrote. He later joined Anthropic.

OpenAI co-founder and chief scientist Ilya Sutskever also left that month. He later founded Safe Superintelligence Inc., whose stated focus is building safe superintelligence.

Daniel Kokotajlo, an OpenAI governance researcher, resigned in 2024 after saying he had lost confidence that the company would handle AGI responsibly. He refused to sign a non-disparagement agreement, forgoing roughly $2 million in vested equity rather than sign, and later became an advocate for protections for employees who warn about AI risks.

Those departures did not establish that OpenAI was ignoring safety. They showed that senior employees had reached serious disagreements over how safety should be weighed against rapid AI development.

The Race Is Now an Industry-Wide Concern

The concern is not limited to people who have left frontier labs. In July, more than 1,100 employees at major AI companies signed a statement asking the U.S. government to help develop technical and governance tools that could deliberately pace the frontier of automated AI development. The statement warned that automated AI research could accelerate beyond humanity’s ability to understand or control the resulting systems.

The signatories included employees from OpenAI, Anthropic, Google, and Meta. The request was not for an immediate halt to AI development. It was for governments to create mechanisms that could buy time if automated AI research began moving faster than safety and oversight could keep up.

Coxon said he ultimately could not accept entering that “endgame” from inside a private company and urged researchers to question whether they should launch increasingly powerful systems without a rigorous understanding of how those systems work.

His departure does not establish that Anthropic or OpenAI will create uncontrollable AI, nor does Hubinger’s estimate amount to an official company forecast. But the record shows a recurring concern among people who have built, studied, or tried to constrain frontier systems: the technology may be advancing toward capabilities that its developers do not yet know how to reliably control.

TERMINAL LAYER

Activate Terminal Layer

Structural analysis of the systems, pressures, and stakeholders behind this story.

FAQ

Frequently Asked Questions

01

Why did pretraining researcher Jacob Coxon leave Anthropic?

Jacob Coxon resigned from Anthropic to protest commercial pressure driving dangerous competitive races toward superintelligence. The 27-year-old researcher spent three years developing core pretraining infrastructure across both OpenAI and Anthropic. Coxon publicly stated that frontier tech companies are deploying increasingly autonomous capabilities without verified alignment controls.
02

Why do researcher departures matter for the artificial intelligence industry?

Exits by core technical staff highlight deep internal skepticism regarding lab safety commitments and commercial release schedules. Prominent researchers like Jan Leike and Ilya Sutskever previously left OpenAI after safety teams lost organizational influence. These recurring resignations erode public trust in corporate self-regulation while accelerating calls for statutory federal oversight.
03

How did Anthropic leadership respond to Jacob Coxon's statements?

Anthropic Alignment Science Lead Evan Hubinger publicly validated Coxon's assessment regarding serious existential risks from advanced systems. Hubinger acknowledged that Anthropic does not yet possess a comprehensive roadmap to solve alignment for recursive superintelligence. Leadership maintains that current models present minimal threat while focusing research efforts on future recursive self-improvement dangers.
04

What systemic risks are former OpenAI and Anthropic employees highlighting?

Former personnel warn that market competition forces commercial laboratories to prioritize product launches over rigorous safety verifications. Former governance researcher Daniel Kokotajlo forfeited roughly $2 million in equity rather than sign non-disparagement confidentiality clauses. Insiders report that rapid recursive software improvement could soon surpass humanity's technical capability to maintain control.
05

What solutions are tech workers proposing to slow the AI race?

More than 1,100 software engineers across Anthropic, OpenAI, Google, and Meta signed a formal governance petition. The coalition asked the United States government to establish technical frameworks that can pace frontier model training. The proposal seeks institutional mechanisms to pause development if automated model iteration outpaces scientific understanding.

You Might Also Like

THE GREY TERMINAL
🛡
Alex Reeve

Alex Reeve is a contributing writer for The Grey Terminal Her articles provide timely insights and analysis across these interconnected industries, including regulatory updates, market trends, token economics, institutional developments, platform innovations, stablecoins, meme coins, policy shifts, and the latest advancements in AI, applications, tools, models, and their broader implications for technology and markets.

The views and opinions expressed by the author in this article are her own and do not necessarily reflect the official position of The Grey Terminal, its management, editors, or affiliates. This content is provided for informational and educational purposes only and does not constitute financial, investment, legal, or tax advice. Readers should conduct their own research and consult qualified professionals before making any decisions related to digital assets, cryptocurrencies, or financial matters. The Grey Terminal and its contributors are not responsible for any losses incurred from reliance on this information.