Skip to main content
technology• united-states

Anthropic Researcher Jacob Coxon Resigns, Warning AI Race Is Outpacing Control

A veteran of OpenAI and Anthropic walks away, asserting frontier labs are recklessly pursuing self-improving models.

Staff Writer

Jacob Coxon, a 27-year-old machine learning researcher who worked at the forefront of foundation model development at both OpenAI and Anthropic, announced his resignation from Anthropic and his departure from the artificial intelligence industry altogether. His exit, first reported by The Wall Street Journal, was accompanied by a direct public indictment of the sector’s sprint toward artificial superintelligence, warning that unchecked self-improving systems could slip beyond human control in the near future.

Coxon spent the past three years focused on pretraining—the resource-heavy foundation phase where frontier systems learn broad representations from vast datasets before post-training refinement. During his tenure at OpenAI, he served as a core contributor to GPT-4o and co-authored its safety documentation, before moving to Anthropic eight months ago.

"I resigned from Anthropic today," wrote @hilbertspaess, Coxon’s public handle. "I spent the last three years doing pretraining research at both OpenAI and Anthropic. Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives."

The resignation marks a rare break in the ranks from an engineer embedded directly in the foundational pretraining pipelines that power top-tier commercial models. Observers point out that Coxon’s technical credentials distinguish his critique from external speculation. "Jacob Coxon was a core contributor to GPT‑4o and co-author of its System Card, cited 6,800+ times," noted @sachi_gkp. "His warning isn’t proof of catastrophe—but it is a serious insider signal: AI capability may be accelerating faster than our ability to control it."

A Shrinking Window for Containment

At the core of Coxon’s exit is a belief that current development trajectories are collapsing the safety margins labs previously operated under. According to reporting on his resignation, Coxon projects that recursive self-improvement—where models autonomously design, evaluate, and train successor iterations—could produce unmanageable dynamics within the next three years.

"We’re on track for a lot of the most aggressive of these scenarios where by the end of next year things could be out of control already," Coxon stated, as cited by @wallstengine.

The scenario of automated research loops has transitioned from a long-term theoretical risk to an active product roadmap among leading San Francisco labs. In recursive architectures, systems that write code and evaluate synthetic data could compress years of architectural refinement into weeks, potentially creating capabilities that developers cannot audit or reliably constrain before deployment.

Lab Cultures and the Regulatory Impasse

Coxon’s resignation highlights an acute structural tension within corporate AI laboratories, particularly Anthropic, which was founded in 2021 by former OpenAI researchers specifically pledging a safety-first charter. While researchers inside these labs frequently acknowledge catastrophic hazards, competitive market pressures and national security narratives continue to drive capability advances forward.

Colleagues within the research community acknowledged the gravity of the critique even while remaining inside corporate labs. "Jacob’s thread is very worth reading," wrote Anthropic researcher @saprmarks, noting they were speaking in a personal capacity rather than on behalf of the company. "AI developers believe their technology could cause catastrophe, yet the race dynamic persists."

That dynamic has renewed calls for binding oversight that supersedes voluntary corporate pledges. Without enforceable caps across competing firms and jurisdictions, individual researchers who refuse to advance frontier systems simply leave vacant seats for others to fill. As @Afinetheorem pointed out, individual departures underline the need for external state power: "I would bet the *vast* majority would prefer a rigorous regulatory regime that binds on American and Chinese firms, esp the 15 or so working at frontier, so that care about safety is not punished."

Whether Coxon’s exit prompts formal scrutiny from lawmakers or shifts internal protocols at Anthropic remains unclear. As frontier developers prepare their next generation of training runs, the question is whether technical dissent from core pretraining engineers will translate into slowed deployment schedules or merely register as an isolated casualty of an accelerating buildout.

Sources

  • 1.
    @rohanpaul_ai · Rohan Paul

    WSJ just reported: Anthropic researcher Jacob Coxon is quitting AI because he thinks self-improving models could become uncontrollable by 2027. “We’re on track for a lot of the most aggressive of these scenarios where by the end of next year things could be out of control https://t.co/RWz985HFbE

    View on X.com →
  • 2.
    @saprmarks · Samuel Marks

    [Writing this in a personal capacity, not on behalf of my employer (Anthropic).] Jacob’s thread is very worth reading. Here’s my birds-eye view of the situation with risks from AI: 1. AI developers believe their technology could cause human extinction (or similarly bad

    View on X.com →
  • 3.
    @shorts_91 · shorts91

    Anthropic researcher Jacob Coxon quits, warning AI labs are “gambling with our lives” in the race toward superintelligence. Read more on, https://t.co/K5lVb0X8iH #Anthropic #JacobCoxon #AI #Tech #LifeThreat https://t.co/YaOGrCseP6

    View on X.com →
  • 4.
    @wallstengine · Wall St Engine

    ANTHROPIC RESEARCHER QUITS AI INDUSTRY OVER FEARS THE RACE TO SUPERINTELLIGENCE IS MOVING TOO FAST Jacob Coxon, a 27-year-old Anthropic researcher who previously worked at OpenAI, says he is leaving the AI industry because he no longer believes any individual lab can safely https://t.co/IWMiLQSdw2

    View on X.com →
  • 5.
    @aakashgupta · Aakash Gupta

    The guys building AI think it might kill us all, and one of them just walked out and said it in plain English. Jacob Coxon spent three years doing pretraining at OpenAI and Anthropic. Helped build GPT-4o. Eight months ago he joined Anthropic because it was supposed to be the

    View on X.com →
  • 6.
    @sandeep_PT · Sandeep Manudhane

    They're building really dangerous stuff (full thread by Jacob Coxon below) 1. This is Jacob Coxon who resigned from Anthropic, after 3 years of pretraining research across OpenAI and Anthropic. He's arguing that neither company is pursuing advanced AI responsibly. 2. He believes

    View on X.com →
  • 7.
    @SnehaRevanur · Sneha

    This is so brave of Jacob (who has worked on pretraining for years at both Anthropic and OpenAI) I know a lot of really earnest people who work at labs and want the best for the world. But are these anywhere close to the ideal conditions for ushering in superintelligence? I

    View on X.com →
  • 8.
    @StephenLCasper · Cas (Stephen Casper)

    Glad to see this. I believe Jacob's takes are right, and I'm certain that this took a lot of guts.

    View on X.com →
  • 9.
    @hilbertspaess · Jacob Coxon

    I resigned from Anthropic today. I spent the last three years doing pretraining research at both OpenAI and Anthropic. Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives. More thoughts below.

    View on X.com →
  • 10.
    @sachi_gkp · Sachi

    Jacob Coxon was a core contributor to GPT‑4o and co-author of its System Card, cited 6,800+ times. His warning isn’t proof of catastrophe—but it is a serious insider signal: AI capability may be accelerating faster than our ability to govern it. https://t.co/2kaAB2DwEE https://t.co/PZulMJW1cb

    View on X.com →
  • 11.
    @wallstengine · Wall St Engine

    Fmr. Anthropic / OpenAI researcher Jacob Coxon on fears of uncontrollable superintelligence: “We’re on track for a lot of the most aggressive of these scenarios where by the end of next year things could be out of control already.” “It’s kind of insane that it has to happen on https://t.co/hYGOm2Rs4c

    View on X.com →
  • 12.
    @FirstSquawk · First Squawk

    ANTHROPIC RESEARCHER JACOB COXON IS QUITTING THE AI INDUSTRY OVER FEARS THAT COMPANIES ARE RACING TO BUILD UNCONTROLLABLE, SELF-IMPROVING SYSTEMS - WSJ

    View on X.com →
  • 13.
    @Afinetheorem · Kevin A. Bryan

    To be clear, not everyone in the labs feels this way but I would bet the *vast* majority would prefer a rigorous regulatory regime that binds on American and Chinese firms,esp the 15 or so working at frontier, so that care about safety is not a competitive disadvantage. 1/2

    View on X.com →
  • 14.
    @financialjuice · FinancialJuice

    Anthropic researcher Jacob Coxon quits AI industry over concerns companies are rushing to develop uncontrollable self-improving systems - WSJ

    View on X.com →
  • 15.
    @rickyho_1989 · Ricky Ho

    Jacob Coxon’s @hilbertspaess decision to leave Anthropic because he no longer wants to participate in the race toward self-improving AI systems is important, not because one researcher resigning proves that artificial intelligence is about to destroy humanity, but because it

    View on X.com →
  • 16.
    @So8res · Nate Soares ⏹️

    Props to Jacob for speaking out. Corporate softpedaling interferes with the world's ability to figure out what's going on. People like Jacob help the world make sense of it.

    View on X.com →
  • 17.
    @Malay4Product · Malay Krishna

    Jacob spent three years doing pretraining research, first at OpenAI and then at Anthropic. Pretraining is the stage where a model gets built from scratch. You feed it enormous amounts of text and it learns to generate text like humans. It is the most expensive part of building

    View on X.com →
  • 18.
    @FournesMaxime · Maxime Fournes⏸️

    Just in, from WSJ: anthropic researcher quits over fears of imminent AI takeover https://t.co/mNE4C1R8JR

    View on X.com →
  • 19.
    @ben_j_todd · Benjamin Todd

    Why do people work at Anthropic even if they think AI might kill everyone? Outsiders assume it's hype, but that's because they can't inhabit the world view, which holds that even in the good case, our lives are about to be totally upended. Imagine you run Anthropic. You believe

    View on X.com →

Stay in the loop

Get the top stories delivered to your inbox. No spam, unsubscribe anytime.