惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

D
DataBreaches.Net
IT之家
IT之家
博客园_首页
博客园 - 【当耐特】
V
V2EX
Apple Machine Learning Research
Apple Machine Learning Research
G
Google Developers Blog
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
Recent Announcements
Recent Announcements
F
Fortinet All Blogs
GbyAI
GbyAI
腾讯CDC
H
Hackread – Cybersecurity News, Data Breaches, AI and More
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
I
InfoQ
H
Help Net Security
T
Tailwind CSS Blog
B
Blog RSS Feed
Martin Fowler
Martin Fowler
人人都是产品经理
人人都是产品经理
The Cloudflare Blog
博客园 - 叶小钗
雷峰网
雷峰网
量子位

Victor Tangermann Archives - Futurism

AI Slop YouTube Channel Glitches Out in a Way So Bizarre That It’s Vaguely Disturbing Double Murder Suspect Asked ChatGPT How to Hide Body in Dumpster The White House Suddenly Seems Pretty Terrified of Anthropic Democrat and Republican Voters United on Key Issue: Hatred of Data Centers NASA’s Moon Landing Schedule Slipping Horrendously Under Trump NASA Fires Up Futuristic Plasma Thruster Designed to Take Us to Mars Toilet Maker Spikes in Value as It Flushes Money Into AI New England Journal of Medicine Retracts Paper Because Photo of Patient’s Insides Was Garbled by AI Eric Trump’s Crypto Company Is Falling Into Total Disaster Gen Z Is Turning Against AI in an Incredible Way If OpenAI Loses This Trial, It Could Effectively Be Eliminated in Its Current Form D4vd Bought Chainsaws and Body Bag on Amazon After Murder, Prosecutors Say, in the Most Grisly Case of Internet Brain Rot We’ve Ever Heard There’s a Hidden Shortcut to Mars, Scientific Paper Finds Reddit Intentionally Breaks Its Mobile Website, Demanding Users Download Its App Instead Uninstalls of ChatGPT Are Spiking at the Worst Time Imaginable for OpenAI Another Military UFO Guy Just Died Sam Altman Caught in What May Be His Most Spectacular Lie Yet Scientists Scan Gigantic Structure Hiding Behind Our Galaxy If You Thought Mark Zuckerberg Was a Pathetic Little Worm Before, Wait Until You Hear About His Latest Move OpenAI in Shambles as IPO Looms Scientists Experimenting With Quantum Effect That Some Fear Could Cause Chain Reaction That Ends Entire Universe Weird Things Happen When You Give AI Agents Money and Let Them Spend It Bitcoin Developers Are Debating a Move That Could Send Crypto Markets Into a Tailspin Video Shows NASA Astronaut Struggling to Walk After Journey Around the Moon Sam Altman Issues Grim Apology Top Medical Journal Publishes Searing Article Warning Against Medical AI New Browser Plugin Adds Typos to Your AI-Generated Emails to Make Them Look Real New AI-Powered Robot Can Destroy Human Champions at Ping Pong Devious New AI Tool “Clones” Software So That the Original Creator Doesn’t Hold a Copyright Over the New Version NASA Planning to Set First-Ever Fire on the Surface of the Moon
Anthropic Was So Concerned About Its New Mythos-Based Mod...
Victor Tangermann · 2026-06-12 · via Victor Tangermann Archives - Futurism

A stylized photo illustration featuring Anthropic co-founder Dario Amodei.

Illustration by Tag Hartman-Simkins / Futurism. Source: Michael M. Santiago / Getty Images; Shutterstock

Sign up to see the future, today

Can’t-miss innovations from the bleeding edge of science and tech

Earlier this year, Anthropic refused to release its Mythos AI model to the public, saying it was simply too dangerous.

At the time, executives claimed the model was capable of punching through powerful cybersecurity safeguards, pointing at researchers who used it to discover thousands of vulnerabilities in widely-used open source code.

Months later, Anthropic was finally ready to go public with the model. On Tuesday, the Dario Amodei-led company announced a Mythos-powered model called Fable 5, which it claims is “safe for general use.”

However, new safeguards quickly frustrated AI researchers, who accused the company of intentionally lobotomizing Fable 5. The backlash was so fierce, Anthropic quickly made adjustments to the policy, as Wired reported on Wednesday, highlighting just how carefully the company is treading.

In its original announcement, Anthropic claimed the safeguards were designed to stop Fable 5 from improving itself, in “new interventions that limit Claude’s effectiveness for requests targeting frontier LLM development.” Just days ahead of the launch, Anthropic released a report on “when AI builds itself,” a trend that “might increase the risks of humans losing control over AI systems.”

However, AI researchers were not impressed by Anthropic hamstringing its latest model’s abilities.

“Anthropic’s latest model will NOT help you if it thinks your ML research/ML engineering is interesting, and/or will secretly degrade its IQ so that the average engineer won’t notice,” AI research firm SemiAnalysis tweeted.

“We are already seeing Anthropic’s latest model’s moderation filters our GPU inference research and programming,” it added.

Other researchers accused Anthropic of using Fable 5 to “shadowban,” or quietly restrict the accounts, of AI researchers. According to the firm’s system card, interventions limiting requests for “frontier LLM development” will “not be visible to the user.”

This last concern, which could’ve effectively sabotaged anybody trying to train competing models by quietly bumping them down to less powerful models without their knowledge, proved controversial enough for Anthropic to change its mind.

“We’re changing Fable 5’s safeguards for frontier LLM development to make them visible,” the company told Wired in a statement. “We made the wrong trade-off and we apologize for not getting the balance right.”

“It felt like Anthropic was saying to the public, ‘We don’t trust anybody else to do AI research,” AI startup Prime Intellect research lead Will Brown told the publication. “We are the only ones who have to do AI research.”

It all comes in the context Anthropic calling for a global freeze on AI advances while discussing the dangers of “recursive self-improvement.” In other words, the company is making a lot of noise about a sci-fi-sounding possibility: that AI will start to rapidly improve itself, potentially escaping the control of its human creators.

Beyond limiting its ability to develop AI tools, Fable 5’s new safeguards also trigger when it encounters requests “related to cybersecurity, biology and chemistry, or distillation.” Distillation is effectively using machine learning to train a “student” model on the behavior and reasoning of a “teacher” model, a practice that has sparked its fair share of controversy.

Anthropic has already publicly griped about large-scale attempts to distill, or “extract” its underlying model — a hypocritical stance given its indiscriminate scraping of rights-protected content on the web to train its AI in the first place.

More on Anthropic: Anthropic Scared, Calls for Global Freeze on AI Advances