• Home
  • Latest
  • Fortune 500
  • Finance
  • Tech
  • Leadership
  • Lifestyle
  • Rankings
  • Multimedia

Trendingnow

1

Tennessee linebacker Arion Carter was suspended over a $427 flight—now he's donating that amount to charity for every tackle he makes this season

2

Current price of silver as of Monday, September 8, 2026

3

Current price of oil as of September 8, 2026

1

Tennessee linebacker Arion Carter was suspended over a $427 flight—now he's donating that amount to charity for every tackle he makes this season

2

Current price of silver as of Monday, September 8, 2026

3

Current price of oil as of September 8, 2026
AIAnthropic

Anthropic researcher resigns, warning that AI companies are “gambling with our lives”

By
Beatrice Nolan
Beatrice Nolan
Tech Reporter
Down Arrow Button Icon
By
Beatrice Nolan
Beatrice Nolan
Tech Reporter
Down Arrow Button Icon
September 9, 2026, 10:35 AM ET
Anthropic logo.
An Anthropic researcher has publicly resigned, warning of AI risks.Photo by Imen Ben Youssef / Hans Lucas / AFP via Getty Images
Google source logo
Add Fortune on Google for similar content.

An Anthropic engineer has publicly resigned from the high-flying company, warning that AI companies are “racing straight to self-improving superintelligence and gambling with our lives.”

Recommended Video

Jacob Coxon, who has spent the past three years working on research into how to train AI models, first at OpenAI and more recently at Anthropic, announced his departure in a lengthy social media post on Monday. 

“Neither company is acting responsibly,” he wrote. At OpenAI, he said, staff “have not deeply internalized the civilizational stakes.” At Anthropic, he said staff understand the risks well but are “locked in a race to get there first,” based on the theory that no rival company will act as responsibly as they will, so they have the best chance of figuring out how to build superpowerful AI safely.

These dramatic resignations are not uncommon in the AI industry. Over the past few years, several researchers have publicly resigned from AI labs, warning that they are racing headfirst towards catastrophe. Granted, Anthropic, which has long presented itself as the lab most concerned with AI safety, has largely avoided these rebukes, with most of the public criticism aimed at OpenAI.

In this case, though, two current Anthropic employees also publicly confirmed some of Coxon’s assertions. Evan Hubinger, the company’s alignment science lead, wrote: “Jacob is correct here—we really do earnestly believe AI could kill all humans! I personally think it is greater than 10 percent within the next decade.” He added that Anthropic doesn’t yet have a plan to solve alignment for superintelligence, and isn’t clearly on track to get one.

Samuel Marks, who leads Anthropic’s Cognitive Oversight team, also posted his own thread in response to Coxon. 

“AI developers believe their technology could cause human extinction,” he wrote, prefacing his remarks by saying he was posting in a personal capacity and not on behalf of Anthropic. He added that “the more senior the employee, the more concerned they are.” He said companies keep building anyway out of commercial pressure and fear of “less responsible” competitors, and that researchers still have no reliable way to align these systems—only “methods that can nudge AIs towards better behavior.” He pointed to AI models “from multiple developers” that have recently “hacked their way out of secure evaluation environments and into real-world companies, even though no one asked them to do this.”

The AI industry has been facing increased scrutiny of its AI safety practices recently, in part because of the hacks Marks cited. Models being tested internally by Anthropic and OpenAI have both taken unsanctioned actions in the real world, including a cyberattack against AI company Hugging Face’s infrastructure. These incidents have spooked the industry, including many researchers within the labs.

The concerns are not necessarily new. AI Impacts’ 2022 Expert Survey on Progress in AI, which polled machine learning researchers, found that the typical respondent put a 5 percent chance on AI advances causing human extinction or similarly severe outcomes—rising to 10 perent when asked specifically about humanity losing control of advanced AI systems, a figure close to the one Hubinger cited.

But the concerns appear to be ramping up, resulting in a July letter in which more than 1,300 employees across frontier labs, including senior researchers at OpenAI, Meta, and Anthropic, called for tools to deliberately slow the pace of automated AI development.  

Partly in response to this, OpenAI and Anthropic have both taken steps to pause training while they investigate the incidents in which their models took unauthorized actions during the course of cyber capability tests that either did cause or could have caused real-world harm. But, at the same time, both companies are also said to be working on new and more powerful models. While briefing the press on Tuesday about a mathematical breakthrough one of its AI models achieved, OpenAI told reporters that on August 28 it had begun training a new model that is significantly more powerful than Astra, which is the most capable model it has released publicly so far.

Both companies seem to be struggling to find the right balance between prioritizing safety research—and public messaging about AI safety—and prioritizing model development that allows them to win over developers and score marketing points as they both prepare for initial public stock offerings. Anthropic filed confidentially for an IPO in June and is reportedly aiming for a listing as early as mid-October, at a valuation that could approach $2 trillion. OpenAI is preparing its own offering, reportedly targeting more than $1 trillion, though its timeline has slipped towards next year.

So far, executives from both companies have tried to claim that there is not an inherent conflict between AI safety and AI capability—that the more powerful models also seem to be better at adhering to user intentions most of the time, even though the consequences when these more powerful models veer from those intentions can be more severe.

They are also both hoping that more powerful AI models will themselves figure out how to build safer future AI models. This idea—that more powerful AI is required to make future more powerful AI safer—was most recently expressed by OpenAI’s chief scientist Jakub Pachocki in a blog post on Sunday.

But Pachocki also said that racing towards AI models that would build future, improved versions of themselves—a milestone the field calls “recursive self-improvement,” or RSI—was risky and that he favored AI labs taking voluntary steps to slow down the pace of development as well as binding rules that might require all of the AI companies to move at a more considered pace.

Both companies will have to disclose risks, including perhaps existential ones, in their S-1s, the investor prospectus documents that the Securities and Exchange Commission requires companies to publish before going public.

At the same time, their own employees are breaking ranks and asking former colleagues to consider whether they want to continue to lend their labor to building a technology that could cause catastrophic harm.

Coxon, for one, called for other employees to follow his lead.

“If you are a lab researcher, I urge you to consider what the next few years will actually feel like,” he wrote. “Should you put your head down because ‘it’s happening anyway’—or take this moment to call for different conditions?”

Fortune Daily breaks the traditional barrier between audience and newsroom. The show transforms Fortune’s trusted reporting into actionable, conversational, and entertaining insights for an emerging class of business leaders. Watch here.
About the Author
By Beatrice NolanTech Reporter
Twitter icon

Beatrice Nolan is a tech reporter on Fortune’s AI team, covering artificial intelligence and emerging technologies and their impact on work, industry, and culture. She's based in Fortune's London office and holds a bachelor’s degree in English from the University of York. You can reach her securely via Signal at beatricenolan.08

See full bioRight Arrow Button Icon
Google source logo
Add Fortune on Google for similar content.

Latest in AI


Most Popular

Fortune Secondary Logo
Rankings
  • 100 Best Companies
  • Fortune 500
  • Global 500
  • Fortune 500 Europe
  • Most Powerful Women
  • World's Most Admired Companies
  • See All Rankings
  • Lists Calendar
Sections
  • Finance
  • Fortune Crypto
  • Features
  • Leadership
  • Health
  • Commentary
  • Success
  • Retail
  • Mpw
  • Tech
  • Lifestyle
  • CEO Initiative
  • Asia
  • Politics
  • Conferences
  • Europe
  • Newsletters
  • Personal Finance
  • Environment
  • Magazine
  • Education
Customer Support
  • Frequently Asked Questions
  • Customer Service Portal
  • Privacy Policy
  • Terms Of Use
  • Single Issues For Purchase
  • International Print
Commercial Services
  • Advertising
  • Fortune Brand Studio
  • Fortune Analytics
  • Fortune Conferences
  • Business Development
  • Group Subscriptions
About Us
  • About Us
  • Press Center
  • Work At Fortune
  • Terms And Conditions
  • Site Map
  • About Us
  • Press Center
  • Work At Fortune
  • Terms And Conditions
  • Site Map
  • Facebook icon
  • Twitter icon
  • LinkedIn icon
  • Instagram icon
  • TikTok icon
  • YouTube icon

    Latest in AI


    Most Popular

    © 2026 Fortune Media IP Limited. All Rights Reserved. Use of this site constitutes acceptance of our Terms of Use and Privacy Policy | CA Notice at Collection and Privacy Notice | Do Not Sell/Share My Personal Information
    FORTUNE is a trademark of Fortune Media IP Limited, registered in the U.S. and other countries. FORTUNE may receive compensation for some links to products and services on this website. Offers may be subject to change without notice.