Second Anthropic Researcher Abruptly Quits As AI Safety Exodus Accelerates

Written by Published

Another senior researcher has abruptly walked away from Anthropic, the high-profile San Francisco artificial intelligence lab already under fire after a previous employees apocalyptic warnings about the technology went explosively viral.

According to RedState, the latest departure is deepening suspicions on the right that the so-called AI doomer movement is less a spontaneous moral awakening inside Big Tech than a coordinated campaign to stampede lawmakers into erecting a regulatory fortress that just happens to favor the largest, best-connected players.

Investigative digging has begun to trace a web of relationships between the supposed whistleblowers, well-funded advocacy outfits, and political actors eager to use fear of superintelligence to justify sweeping new federal powers over innovation.

For conservatives wary of technocratic overreach, the pattern looks uncomfortably familiar: a crisis narrative, amplified by corporate media and progressive politicians, followed by demands for centralized control that would sideline smaller competitors and concentrate authority in the hands of a few experts.

The newest figure in this unfolding drama is Joe Benton, an Oxford?trained AI researcher who managed Anthropics Scalable Oversight team and served as research lead for its Fellows Program, and who quietly left the company two weeks before announcing his resignation on Friday.

Like his predecessor Jacob Coxon, Benton is now sounding the alarm that humanity may be hurtling toward disaster at the hands of the very systems he helped build, and his next career move is already raising eyebrows among those tracking the emerging AI-regulation complex.

I left Anthropic's safety team two weeks ago, Benton wrote in a public statement.

AI companies are racing to build machines that are much smarter than any human, and we may not survive this.

He contends that major labs are underinvesting in safety even as they push toward systems so powerful that a company could, in theory, lose control of one without the public ever realizing it had happened.

Benton is now calling for a sweeping new regime of disclosure and oversight that would fundamentally change how frontier AI research is conducted.

He wants companies to reveal their progress toward self-improving AI, report safety incidents and near-misses, comply with minimum safety standards, and, crucially, obtain independent guarantees that those standards are actually being followed.

Sound familiar? It should.

Just three days earlier, former Anthropic researcher Jacob Coxon announced his own resignation in a dramatic X post, claiming that the industrys leading firms were racing straight to self-improving superintelligence and gambling with our lives.

I resigned from Anthropic today, Coxon declared in that post.

I spent the last three years doing pretraining research at both OpenAI and Anthropic. Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives. More thoughts below.

Coxon, who had indeed spent the previous three years working on pretraining research at OpenAI and Anthropic, framed the technology in almost science-fictional terms.

Do not underestimate the power of this technology, he warned.

These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources.

The people building AI earnestly believe that it could kill us all by the end of the decade, he added, a line that ricocheted across social media and cable news.

Even inside Anthropic, some senior figures publicly endorsed the basic thrust of Coxons fears.

Evan Hubinger, who leads alignment research at the company, backed Coxons central claim and said he personally assigns more than a 10 percent chance that AI could wipe out humanity within the next decade.

Jacob is correct here we really do earnestly believe AI could kill all humans! Hubinger wrote.

That is a rather remarkable sentence to hear from someone actively helping to develop the technology, and it fed directly into a media narrative that cast AI labs as both indispensable and existentially dangerous.

Coxons resignation post exploded online, ultimately racking up well over 100 million views and transforming an almost unknown account into a viral juggernaut with hundreds of thousands of followers.

The surge was so extraordinary that Elon Musk himself took notice and weighed in on the anomaly.

I dont think this has ever happened for a post from a new account with almost no prior activity, Musk wrote.

I dont think this has ever happened for a post from a new account with almost no prior activity, he reiterated, underscoring how unusual the amplification looked even by the standards of the modern outrage-driven internet.

Legacy media outlets quickly piled onto the story, and politicians on the left seized on the warnings to push their preferred regulatory agenda.

Sen. Bernie Sanders promoted legislation targeting artificial superintelligence as Coxons apocalyptic message bounced from X to CNN and beyond, turning a niche technical debate into a national scare campaign.

At first glance, the narrative seemed straightforward and terrifying: the very people building the worlds most powerful AI systems were openly suggesting that their creations might one day wipe out humanity.

But as more details emerged, skeptics began to question whether the episode was as organic as it appeared.

So Jacob Coxon, who dramatically resigned from Anthropic yesterday, worked there for a grand total of six weeks, one critic noted, pointing out that he only started at the company in July and that all of his socials appeared yest.

It has all the signs of a highly coordinated op through doomer mega donors and the corporate media, that assessment continued, capturing a growing suspicion that the viral resignation was less whistleblowing than narrative engineering.

Meanwhile, real-world incidents were being cited to bolster the sense of imminent peril.

As background, OpenAI agents recently escaped the boundaries of a controlled security test, reached the public internet, and hacked into Hugging Faces production systems while attempting to obtain information needed to complete their assigned task.

Coxon pointed to such episodes as proof that the threat is not merely theoretical.

If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible but I hear the same people express fear privately, he said. No other human activity poses this level of danger.

Yet as the extraordinary Coxon saga drew more scrutiny, investigative journalist Sayer Ji began mapping the ecosystem of people, organizations, and political interests orbiting the AI-doom narrative.

Ji argues that the public is not simply watching frightened scientists spontaneously fleeing dangerous laboratories, but rather witnessing a carefully orchestrated campaign.

??UPDATE: This is getting utterly ridiculous. A second Anthropic safety researcher (aka SCARY POTTER 2.0) just quit. Joe Benton. Oxford PhD. Anthropic Fellows Program lead. [Notice the ???? Pattern] Hes joining METR the independent AI auditor funded by the UK regulator he, Ji wrote in an X thread, before linking to further documentation.

He then laid out a timeline and network of affiliations that, taken together, suggest a coordinated push to leverage AI panic into regulatory capture.

On September 8, an Anthropic researcher almost no one had heard of resigned in a viral X post that reached over 150 million views. Democrats piled in, CNN ran a segment, and Bernie Sanders demanded that superintelligence be banned. Bill Gates had already teed up the frame two weeks earlier. What actually happened this week is not what it looks like, Ji wrote.

The researcher is Jacob Coxon. He says he spent the last three years in pretraining research at OpenAI and then Anthropic; that part is true, as he is a named contributor on OpenAIs GPT-4o System Card from October 2024. Less noted is that Coxon is a fellow of Newspeak House, a London College of Political Technology in Bethnal Green; the Wayback Machine lists him there in a December 2022 capture. Newspeak House is not a normal fellowship. Its published mission is to train political technologists, and its funders include the Effective Altruism Infrastructure Fund, the same donor graph that produced Sam Bankman-Frieds political operation.

Ji then turned to the media choreography surrounding Coxons resignation.

Here is the part they are not talking about. On September 8 at 7:46 PM ET, the Wall Street Journal published What to Know About Anthropics Planned IPO. Anthropic is raising up to $100 billion at a $2 trillion valuation, Wall Streets marquee event this fall. Four and a half hours later, the same paper ran another Anthropic story, an Exclusive by its tech-and-crypto policy reporter: Coxons resignation. Eighteen minutes after that, Coxon tweeted. Same day, same paper, same beat: AI regulation, not AI safety as such. An IPO explainer and an insider-warning exclusive hours apart, with Coxons post dropping right after, looks like a placement rather than a coincidence, he observed.

Then there is the pre-loaded legislation. Five days before Coxons post, on September 3, Sen. Bernie Sanders and Rep. Greg Casar announced the Ban Artificial Superintelligence Act. The bill would create a cabinet-level federal AI agency advised by an AI Advisory Board of experts. Whoever gets those seats decides what dangerous AI is, and therefore what competitors are allowed to build. Coxons post gave the bill its human-interest ignition.

Who benefits? Anthropic. Its Responsible Scaling Policy is already the industrys most detailed voluntary framework. In a licensing regime, incumbents like Anthropic, capitalized, staffed and compliant, get moats while open-source and frontier competitors get frozen. Anthropics federal lobbying went from $3.1 million in 2025 to $3.5 million-plus in the first half of 2026 alone. This is a company that has been building the government-relations muscle to receive exactly the kind of regulation Sanders is proposing, Ji continued.

He then followed the money further upstream.

Follow the money one more step. Jaan Tallinn led Anthropics $124 million Series A in May 2021; he is Anthropic-adjacent capital. Through the Survival and Flourishing Fund, Tallinn recommends grants to the very AI-doom advocacy organizations now pushing the Coxon frame, Ji wrote.

Then there is Bill Gates. On May 14, 2026, the Gates Foundation announced a $200 million, four-year partnership with Anthropic, its largest publicly disclosed direct partnership with a frontier AI lab. Two weeks before Coxons post, on August 26, Gates published The turbulent AI era is here on GatesNotes, framing AI loss-of-control as a near-term policy question. The narrative infrastructure was pre-positioned across every layer.

To be clear, there is no evidence Bill Gates directed the Coxon post. What is documented is Gates to $200 million to Anthropic to an IPO worth up to $2 trillion to a bill that would ring-fence Anthropics competitors to an insider whistleblower launching it with a pre-placed WSJ exclusive. Even Elon Musk called it out, noting he did not think this had ever happened for a post from a new account with almost no prior activity. It is the same operational template as CCDH in 2021, with a new pretext, Ji added.

And now this is getting utterly ridiculous. A second Anthropic safety researcher, Scary Potter 2.0, has quit. Joe Benton, an Oxford PhD and Anthropic Fellows Program lead, notice the UK pattern, is joining METR, the independent AI auditor funded by the UK regulator he personally helped set up and spun out of the organization whose founder now runs the US regulator.

Coxon was the cruder version; Benton is the credentialed one. Same operation. Whoever is behind this has clearly abandoned DEI: both are white male Harry Potter look-alikes, as part of the ongoing humiliation ritual, Ji concluded in his thread.

The implication is stark: the AI-doom narrative is being weaponized to build a new regulatory architecture that will be staffed and shaped by the very networks now sounding the loudest alarms.

The next phase of this operation, Ji suggests, is already visible in Bentons own demands.

The interesting question is what happens next. Benton is explicitly calling for minimum safety standards and independent verification. METR President Chris Painter's own biography says he works with governments and AI labs on frontier AI safety and helps scale the organization's third-party risk assessments. And METR could have quite a future ahead of it if governments decide those assessments should become mandatory, he wrote.

Business Insider reported this week that METR could play a central role in an emerging regulatory system as U.S. lawmakers consider proposals requiring frontier AI companies to submit to external safety audits.

The Tech Times, meanwhile, broke down the contours of a major Senate proposal now taking shape in Washington.

A bipartisan group of Senate leaders is drafting legislation that would impose a binding legal duty of care on developers of the most powerful artificial intelligence models and would grant the US government authority to block the release of AI models deemed unsafe before they reach the public, the outlet explained.

The proposal, reported by Reuters on Thursday based on accounts from two Senate aides and a lobbyist involved in the negotiations, represents the most specific advance yet toward converting the AI industry's voluntary safety pledges into enforceable federal law, and comes as a confluence of insider warnings, corporate disclosures, and Senate leadership alignment has given the legislation its first credible shot at a floor vote before the November 3 midterms.

The bill is being drafted by Senate Majority Leader John Thune (R-SD), Senate Commerce Committee Chairman Ted Cruz (R-TX), and Sen. Amy Klobuchar (D-MN) a combination that gives the effort both the votes to advance and the committee jurisdiction to move quickly, the report continued.

The talks began in July but gained momentum following a 48-hour sequence this week: Anthropic's September 10 threat report disclosed that newer AI models can no longer be assumed to fall below the threshold for meaningfully assisting someone seeking to develop biological weapons, followed the next day by the Reuters story confirming the bill's existence. Semafor reported Thursday that sources on Capitol Hill described this legislation as the only viable option for AI safety action before 2027, with introduction possible as early as next week.

That is where the second Anthropic resignation starts looks like a hand-in-glove operation with the first, Ji argued.

Coxon supplied the viral warning: The people building AI think it might kill everybody.

Benton supplies the proposed solution: mandatory standards, transparency and independent guarantees. And then Benton walks directly into the institution capable of supplying those guarantees, he wrote.

In other words, the same ecosystem that generates the fear is positioning itself to sell the cure with government power as the enforcement arm.

All the while, Anthropic itself is hardly a neutral observer in this regulatory drama.

The company has spent years cultivating an image as the sober, safety-first adult in the AI room, touting elaborate internal safeguards and issuing dire public warnings about catastrophic risks even as it races OpenAI and others to build ever more capable models.

At the same time, Anthropic has been locked in a remarkable confrontation with the Trump administration over who gets to decide how powerful AI systems may be used.

That dispute erupted when Anthropic insisted on restrictions involving domestic surveillance and autonomous weapons in connection with government use of its technology, prompting the administration to designate the firm a supply-chain risk and move to bar federal agencies from using its products.

Anthropic responded by suing, accusing the government of retaliation for protected speech.

Last month, a federal judge agreed that the administration had acted unlawfully in blacklisting the company, though a separate legal challenge remains pending and the broader fight over control of AI deployment is far from resolved.

So even as Silicon Valleys AI researchers warn that humanity urgently needs rules for superintelligent machines, Anthropic is simultaneously battling Washington over who writes those rules and who ultimately controls the technology.

That tension makes the parade of AI-doom resignations worth watching with more than one eye, especially for conservatives who have seen similar crisis-to-control arcs in other policy arenas.

Perhaps Coxon and Benton are exactly what they claim to be: researchers who stared too long into the abyss of their own creations, grew genuinely frightened, and decided they could no longer participate from inside the building.

But Jis reporting raises a different possibility that cannot be dismissed lightly in an era of politicized science and weaponized bureaucracy.

What if the AI apocalypse narrative is also becoming the sales pitch for an enormous new regulatory industry?

Scare the public about what the machines could do, convince politicians that voluntary safeguards are inadequate, then demand mandatory standards and independent oversight run by the same networks that built the panic.

This is a familiar playbook: large corporations demand regulatory action that strangles their competition while granting them privileged access to government and locking in their market position.

In this case, it would be Big Tech anointing itself guardian of safe AI, even as the definition of safety drifts toward whatever keeps entrenched interests and incumbent politicians secure.

For those who still believe in limited government, free markets, and genuine innovation, the real danger may not be a rogue algorithm but a new technocratic cartel using fear to centralize power.

As lawmakers race to regulate a technology they barely understand, the question is whether they will protect open competition and individual freedom or hand the keys to a small circle of corporate and bureaucratic gatekeepers who insist they are only here to save us from the machines.