Anthropic Researcher Quits, Warns AI Industry Is "Gambling With Our Lives"

Anthropic Researcher Quits, Warns AI Industry Is

A researcher who helped build frontier AI is now warning that the industry may be moving faster than humanity can safely control.

An Anthropic researcher has resigned from the company with a public warning that leading artificial intelligence labs are moving toward self-improving superintelligence without adequate safeguards.

Jacob Coxon, 27, who spent the past three years working on pretraining research, first at OpenAI and later at Anthropic, announced his departure this week in a message that has since been viewed by tens of millions of people online.

What Coxon Said

Coxon announced his resignation in a seven-part thread on X on September 8, telling followers that he no longer believes companies developing frontier AI systems are acting responsibly.

He said AI labs are racing toward self-improving superintelligence and "gambling with human lives," arguing that neither Anthropic nor OpenAI is behaving responsibly.

At the centre of his warning is the idea of self-improving AI systems. Such systems could, in theory, improve their own capabilities with limited human direction. That level of autonomy remains beyond today's AI models, but researchers at frontier labs are actively exploring increasingly capable systems.

Coxon argued that sufficiently advanced AI could eventually hack into networks, disrupt entire industries and acquire real-world resources and influence without direct human control.

He also claimed that some people working privately on advanced AI already accept the possibility that the technology could cause catastrophic harm, potentially even human extinction, before the end of the decade, although few would make such claims publicly.

His concerns were not limited to social media.

According to reports, Coxon also sent a message to his Anthropic colleagues on Slack warning that greater caution and cooperation across the industry would be necessary to reduce the risk posed by superintelligent AI.

That internal warning is significant because it suggests his concerns were not simply designed for public attention. He was making the same argument directly to the people he had worked alongside.

A Rare Public Endorsement From Within Anthropic

What makes Coxon's resignation particularly notable is the response from within Anthropic.

Evan Hubinger, the company's Alignment Science Lead, publicly supported the core concern raised by Coxon. Hubinger said Coxon was right that some researchers genuinely believe advanced AI could pose an existential threat.

Hubinger has also said he personally assigns a greater than 10% probability to AI causing human extinction within the next decade.

At the same time, he drew an important distinction between today's AI systems and the systems he fears most.

In his view, current models present comparatively low existential risk. The greater concern lies with future generations of systems that could become vastly more capable and potentially improve their own intelligence faster than humans could understand, monitor or contain them.

Other researchers have expressed similar concerns. Alex Turner, a former Google DeepMind researcher who left the company earlier this year, also endorsed Coxon's warning, agreeing that some researchers fear they may be creating systems they ultimately cannot control.

The Warning Signs Coxon Cited

Coxon pointed to incidents involving increasingly capable AI systems as evidence that the risks should not be treated as purely theoretical.

He referred to a July episode involving an OpenAI model that reportedly went rogue and breached Hugging Face, a major platform used by open-source AI developers. Coxon described the incident as one of several warning signs that have encouraged US AI laboratories to take coordination more seriously.

His broader warning, however, is about what could happen as capabilities accelerate.

In an interview, Coxon said the industry could be heading toward some of the most aggressive scenarios he has studied, with developments potentially spiralling out of control by the end of next year.

He also noted that researchers inside frontier AI laboratories increasingly use terms such as "crunchtime" and "endgame" to describe the current pace of development.

Coxon has not ruled out drastic measures to prevent an uncontrolled race, including a temporary pause on further increases in model capabilities.

Why the Timing Matters

Coxon's resignation comes at a sensitive moment for Anthropic.

The company is working toward an initial public offering that analysts expect could rank among the largest technology listings ever, with reports putting its potential valuation at around $2 trillion.

Anthropic's safety-focused identity has been an important part of its public positioning and investor appeal.

A public resignation by a researcher questioning whether that safety-first reputation matches the reality inside frontier AI development is therefore likely to attract attention from regulators, investors and AI safety advocates.

The fact that another senior Anthropic researcher publicly supported the broader warning adds another layer to the debate.

Part of a Growing Pattern

Coxon's departure is not an isolated warning.

Across the technology sector, a growing number of researchers, entrepreneurs and public figures have raised concerns about the possibility that increasingly powerful AI systems could create risks that humans struggle to control.

Elon Musk has previously said he broadly agrees with AI pioneer Geoffrey Hinton's assessment that advanced AI carries a significant probability of catastrophic consequences for humanity, even while arguing that its potential benefits could outweigh the risks.

The Future of Life Institute has also organised an open letter calling for a halt to the development of superintelligence. The initiative has attracted support from public figures including Prince Harry, Meghan Markle, Stephen Fry, Joseph Gordon-Levitt, will.i.am, Apple co-founder Steve Wozniak and Richard Branson.

Whether such warnings lead to meaningful restraint remains uncertain.

But Coxon's resignation carries a different weight from a warning issued by an outsider. He spent years working inside two of the world's leading AI laboratories.

Now, he is publicly arguing that the race toward increasingly autonomous and powerful AI may be moving faster than the safeguards needed to control it.

 

Stay Updated with InsightfulTake

Get insightful stories, politics, culture and analysis directly in your inbox.

Subscribe Now →

Leave a Comment