• Home
  • Latest
  • Fortune 500
  • Finance
  • Tech
  • Leadership
  • Lifestyle
  • Rankings
  • Multimedia

Trendingnow

1

Mark Cuban says he has the solution to growing income inequality, and it's to reward every employee—from CEO to janitor—with company stock

2

‘I want to die broke’: Billionaire philanthropist Denny Sanford dies after giving away $4 billion

3

'Dr. Doom' Nouriel Roubini says we're headed for universal basic income or 'some form of socialism' as AI revolutionizes work—He calls that optimistic

1

Mark Cuban says he has the solution to growing income inequality, and it's to reward every employee—from CEO to janitor—with company stock

2

‘I want to die broke’: Billionaire philanthropist Denny Sanford dies after giving away $4 billion

3

'Dr. Doom' Nouriel Roubini says we're headed for universal basic income or 'some form of socialism' as AI revolutionizes work—He calls that optimistic
AIAnthropic

Anthropic walks back covert capability limits on Claude Fable 5 after being accused of ‘secret sabotage’ by AI researchers and developers

Sharon Goldman
By
Sharon Goldman
Sharon Goldman
AI Reporter
Down Arrow Button Icon
Sharon Goldman
By
Sharon Goldman
Sharon Goldman
AI Reporter
Down Arrow Button Icon
June 10, 2026, 1:43 PM ET
Samuel Boivin—NurPhoto/Getty Images
Add Fortune on Google for similar content.

Anthropic has walked back a policy that covertly limited the ability of its new Mythos-tier model, Claude Fable 5, to be used for AI research.

Recommended Video

When Anthropic made its first Mythos-tier model available to the general public yesterday, Fortune reported it was a “considerable step” for the lab, coming just over a week after the company confidentially filed for IPO paperwork. It had initially deemed Mythos-class models too dangerous to release, citing their significantly enhanced ability to identify software vulnerabilities, but said it was now confident new guardrails in Claude Fable 5 are enough to ensure these dangerous skills don’t fall into the wrong hands.

Just hours after the model’s release, however, major backlash from AI researchers, developers, and policy experts began brewing on social media. The pushback centered around a paragraph buried in Claude Fable 5’s 319-page system card—a document that offers detailed safety disclosures—which revealed that Fable would quietly downgrade its own responses when it detected requests related to cutting-edge AI development work, such as building the infrastructure used to train large AI models.

In practice, that means a user could ask Fable for help, receive a deliberately weakened answer, but not know the model was holding anything back. Critics made it clear they felt this undermined a basic expectation that a tool would either do what it was asked or tell the user it wouldn’t.

Unlike Fable’s other restrictions, such as around cybersecurity and biology, which openly redirect users to a less powerful model with a visible notification, the system card emphasized that this is “not visible to the user.” The model still responds, but uses “interventions to limit Claude’s effectiveness” without telling the user it’s doing so.

Anthropic estimated the restrictions would affect roughly 0.03% of traffic. But it also defended its effort by saying “enforcing this restriction through our safeguards avoids accelerating the actors most willing to violate these terms.” 

Later, the company told Fortune it had decided to change Fable 5’s safeguards for frontier LLM development to make them visible.

“We made the wrong tradeoff, and we apologize for not getting the balance right,” an Anthropic spokesperson said. “Building these safeguards is a complex technical challenge: users may experience more false positives as we refine these classifiers to respond to new threats. We are working to reduce these as fast as possible.”

It’s not the first time users have accused Anthropic of being less than fully transparent. Earlier this year, the company faced a wave of complaints after quietly rolling out changes to its Claude Code tool that developers said degraded the tool’s performance.

Pushback from AI community

A wide swath of the AI community pushed back sharply—including open-source researchers critical of Anthropic’s closed policies, as well as AI safety experts who typically align with Anthropic.

“To have my access to the cutting edge models for my work rug pulled in an under the table fashion is appalling,” wrote Nathan Lambert, an open-model researcher who most recently led work at AI2. “To me this paints Anthropic clearly as anti-science, and therefore anti-progress and anti-safety.” 

Dean Ball, a senior fellow at the Foundation for American Innovation who previously served as senior policy advisor at the White House Office of Science and Technology Policy, wrote that Anthropic’s “secret sabotage” safety policy “massively and profoundly raises the status of the argument that AI safety has been hype to justify monopolistic behavior by labs.” 

And Jeremy Howard, head of nonprofit research group Fast AI, wrote that “Anthropic has chosen the opposite of the safe path: they are allowing themselves, the current top lab, to use their top model for frontier AI research. They’ve said they’ll sabotage others who try. This means the AI frontier advances, & power imbalance increases.” 

Even former Anthropic employees joined in. Behnam Neyshabur, who previously co-led Anthropic’s effort to develop an AI scientist, posted on X saying: “Working on AI for cancer? Sorry, I can’t help you. Working on AI for Alzheimer’s Disease? Sorry, I’m becoming a bit dumb when it comes to the AI part of it.” In another post, he added: “I’ve argued for the last eight months that this was the direction things were heading. In my view, concentrating these capabilities fundamentally slows scientific and technological progress and is net negative for humanity.”

Not all prominent AI voices weighed in with criticism, however. Ethan Mollick, an associate professor at Wharton studying AI, innovation, and entrepreneurship, did not focus on the restrictions, writing in a blog post that Claude Fable 5 “outperformed basically every other public model I have used by a considerable margin.” 

Former OpenAI cofounder and Tesla AI director Andrej Karpathy, who announced he had joined Anthropic last month, called Claude Fable 5 a “super exciting release” on X and said it is a “major-version-bump-deserving step change forward.” He did, however, point out that the model “still has quirks that people will run into and the safeguards are configured to be a little too trigger-happy for launch, which can hopefully be tuned over time.”  

Anthropic says it wants to make models accessible and safe

Before the release, Anthropic seemed to gird itself for backlash, though it did not specifically address potential blowback regarding the research restrictions. In an interview with Fortune yesterday, Dianne Na Penn, Anthropic’s head of product management, research, and labs, said that the new model was able to produce frontier performance that was 10 to 20 points more than its previous model, Opus 4.8 or other frontier models.

“I think generally being able to do that, at the same time having the right guardrails in place to make it accessible, and generally in a safe manner, I think that’s probably the main thing that I want folks to take away,” she said. “We’re raising the bar on the intelligence of the models, and at the same time, we are pushing the frontier in a safe manner.” 

She added that Anthropic recognized that some benign requests would initially be blocked. “We’re working actively on making those safeguard improvements post-launch, but we wanted to make the model accessible generally in a safe manner as soon as we could.”

Subscribe to Fortune Gulf Brief. Every Tuesday, this new newsletter delivers clear-eyed, authoritative intelligence on the deals, decisions, policies, and power shifts shaping one of the world’s most consequential regions, written for the people who need to act on it. Sign up here.
About the Author
Sharon Goldman
By Sharon GoldmanAI Reporter
LinkedIn icon

Sharon Goldman is an AI reporter at Fortune and co-authors Eye on AI, Fortune’s flagship AI newsletter. She has written about digital and enterprise tech for over a decade.

See full bioRight Arrow Button Icon
Add Fortune on Google for similar content.

Latest in AI

Finance
Lorem ipsum dolor sit amet, consectetur adipiscing elit, sed do eiusmod tempor incididunt ut labore et dolore magna aliqua. Ut enim ad minim veniam
By Fortune Editors
October 20, 2025
Finance
Lorem ipsum dolor sit amet, consectetur adipiscing elit, sed do eiusmod tempor incididunt ut labore et dolore magna aliqua. Ut enim ad minim veniam
By Fortune Editors
October 20, 2025
Finance
Lorem ipsum dolor sit amet, consectetur adipiscing elit, sed do eiusmod tempor incididunt ut labore et dolore magna aliqua. Ut enim ad minim veniam
By Fortune Editors
October 20, 2025
Finance
Lorem ipsum dolor sit amet, consectetur adipiscing elit, sed do eiusmod tempor incididunt ut labore et dolore magna aliqua. Ut enim ad minim veniam
By Fortune Editors
October 20, 2025
Finance
Lorem ipsum dolor sit amet, consectetur adipiscing elit, sed do eiusmod tempor incididunt ut labore et dolore magna aliqua. Ut enim ad minim veniam
By Fortune Editors
October 20, 2025
Finance
Lorem ipsum dolor sit amet, consectetur adipiscing elit, sed do eiusmod tempor incididunt ut labore et dolore magna aliqua. Ut enim ad minim veniam
By Fortune Editors
October 20, 2025

Most Popular

Finance
Lorem ipsum dolor sit amet, consectetur adipiscing elit, sed do eiusmod tempor incididunt ut labore et dolore magna aliqua. Ut enim ad minim veniam
By Fortune Editors
October 20, 2025
Finance
Lorem ipsum dolor sit amet, consectetur adipiscing elit, sed do eiusmod tempor incididunt ut labore et dolore magna aliqua. Ut enim ad minim veniam
By Fortune Editors
October 20, 2025
Finance
Lorem ipsum dolor sit amet, consectetur adipiscing elit, sed do eiusmod tempor incididunt ut labore et dolore magna aliqua. Ut enim ad minim veniam
By Fortune Editors
October 20, 2025
Finance
Lorem ipsum dolor sit amet, consectetur adipiscing elit, sed do eiusmod tempor incididunt ut labore et dolore magna aliqua. Ut enim ad minim veniam
By Fortune Editors
October 20, 2025
Finance
Lorem ipsum dolor sit amet, consectetur adipiscing elit, sed do eiusmod tempor incididunt ut labore et dolore magna aliqua. Ut enim ad minim veniam
By Fortune Editors
October 20, 2025
Finance
Lorem ipsum dolor sit amet, consectetur adipiscing elit, sed do eiusmod tempor incididunt ut labore et dolore magna aliqua. Ut enim ad minim veniam
By Fortune Editors
October 20, 2025
Fortune Secondary Logo
Rankings
  • 100 Best Companies
  • Fortune 500
  • Global 500
  • Fortune 500 Europe
  • Most Powerful Women
  • World's Most Admired Companies
  • See All Rankings
  • Lists Calendar
Sections
  • Finance
  • Fortune Crypto
  • Features
  • Leadership
  • Health
  • Commentary
  • Success
  • Retail
  • Mpw
  • Tech
  • Lifestyle
  • CEO Initiative
  • Asia
  • Politics
  • Conferences
  • Europe
  • Newsletters
  • Personal Finance
  • Environment
  • Magazine
  • Education
Customer Support
  • Frequently Asked Questions
  • Customer Service Portal
  • Privacy Policy
  • Terms Of Use
  • Single Issues For Purchase
  • International Print
Commercial Services
  • Advertising
  • Fortune Brand Studio
  • Fortune Analytics
  • Fortune Conferences
  • Business Development
  • Group Subscriptions
About Us
  • About Us
  • Press Center
  • Work At Fortune
  • Terms And Conditions
  • Site Map
  • About Us
  • Press Center
  • Work At Fortune
  • Terms And Conditions
  • Site Map
  • Facebook icon
  • Twitter icon
  • LinkedIn icon
  • Instagram icon
  • Pinterest icon

Latest in AI

LaLiga president says tech is already making the fan experience better, and it could solve football’s money problem too
AISports
LaLiga president says tech is already making the fan experience better, and it could solve football’s money problem too
By Catherina GioinoJuly 20, 2026
7 hours ago
Photo of Hugging Face cofounder and CEO Clement Delangue.
CybersecurityAI agents
Hugging Face says it resorted to a Chinese AI model to battle a fully autonomous cyberattack because U.S. model guardrails stymied its defense
By Emily ForliniJuly 20, 2026
9 hours ago
Sumeet Agrawal is VP of Product Management (Data, AI Governance & Context Engineering for Agentic Systems) at Salesforce.
CommentaryAI agents
Salesforce VP on the leaky AI pipeline: why cheaper tokens won’t fix enterprise AI
By Sumeet AgrawalJuly 20, 2026
14 hours ago
Artificial Intelligence technology and futuristic technology transformation
C-SuiteCFO Daily
How CFOs can tell if AI is actually creating value
By Sheryl EstradaJuly 20, 2026
15 hours ago
ryan
CommentaryLeadership
I’ve interviewed 700 leaders. The ones outsourcing their thinking to AI lose their best people first
By Ryan HawkJuly 20, 2026
16 hours ago
Korea’s AI-heavy market now sets the tone for global stocks
InvestingSouth Korea
Korea’s AI-heavy market now sets the tone for global stocks
By Winnie Hsu, Momoka Yokoyama, Masaki Kondo and BloombergJuly 19, 2026
2 days ago

Most Popular

Mark Cuban says he has the solution to growing income inequality, and it's to reward every employee—from CEO to janitor—with company stock
Success
Mark Cuban says he has the solution to growing income inequality, and it's to reward every employee—from CEO to janitor—with company stock
By Sasha RogelbergJuly 20, 2026
10 hours ago
‘I want to die broke’: Billionaire philanthropist Denny Sanford dies after giving away $4 billion
Success
‘I want to die broke’: Billionaire philanthropist Denny Sanford dies after giving away $4 billion
By Sydney LakeJuly 20, 2026
14 hours ago
'Dr. Doom' Nouriel Roubini says we're headed for universal basic income or 'some form of socialism' as AI revolutionizes work—He calls that optimistic
AI
'Dr. Doom' Nouriel Roubini says we're headed for universal basic income or 'some form of socialism' as AI revolutionizes work—He calls that optimistic
By Jason MaJuly 18, 2026
2 days ago
Warren Buffett says his $147 billion investing career was an accident: ‘I may be one of the 10 luckiest in the world’
Future of Work
Warren Buffett says his $147 billion investing career was an accident: ‘I may be one of the 10 luckiest in the world’
By Sarah GlodekJuly 20, 2026
21 hours ago
Current price of silver as of Monday, July 20, 2026
Personal Finance
Current price of silver as of Monday, July 20, 2026
By Joseph HostetlerJuly 20, 2026
18 hours ago
Power companies are using eminent domain to seize land for data centers as 70% of Americans say not in my backyard
AI
Power companies are using eminent domain to seize land for data centers as 70% of Americans say not in my backyard
By Aaron Walayat and The ConversationJuly 19, 2026
2 days ago

© 2026 Fortune Media IP Limited. All Rights Reserved. Use of this site constitutes acceptance of our Terms of Use and Privacy Policy | CA Notice at Collection and Privacy Notice | Do Not Sell/Share My Personal Information
FORTUNE is a trademark of Fortune Media IP Limited, registered in the U.S. and other countries. FORTUNE may receive compensation for some links to products and services on this website. Offers may be subject to change without notice.