惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

云风的 BLOG
云风的 BLOG
博客园 - 三生石上(FineUI控件)
WordPress大学
WordPress大学
F
Fortinet All Blogs
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
博客园 - 叶小钗
爱范儿
爱范儿
美团技术团队
H
Hackread – Cybersecurity News, Data Breaches, AI and More
有赞技术团队
有赞技术团队
博客园_首页
T
The Blog of Author Tim Ferriss
T
Tailwind CSS Blog
V
Visual Studio Blog
Jina AI
Jina AI
博客园 - Franky
量子位
MongoDB | Blog
MongoDB | Blog
L
LangChain Blog
Apple Machine Learning Research
Apple Machine Learning Research
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
U
Unit 42
aimingoo的专栏
aimingoo的专栏
M
MIT News - Artificial intelligence

Futurism

Frontier AI Models Giving Specific, Actionable Instructions to Perpetrate Bioterror Attack AI Slop YouTube Channel Glitches Out in a Way So Bizarre That It’s Vaguely Disturbing Double Murder Suspect Asked ChatGPT How to Hide Body in Dumpster An Elegant Solution to AI Slop: Tax It, and Use the Resulting Billions of Dollars to Fund Cultural Institutions, Artists, and Researchers The White House Suddenly Seems Pretty Terrified of Anthropic Democrat and Republican Voters United on Key Issue: Hatred of Data Centers Chinese Court Rules That a Worker Cannot Be Replaced by AI Toilet Maker Spikes in Value as It Flushes Money Into AI New England Journal of Medicine Retracts Paper Because Photo of Patient’s Insides Was Garbled by AI Gen Z Is Turning Against AI in an Incredible Way If OpenAI Loses This Trial, It Could Effectively Be Eliminated in Its Current Form AI Spy Cameras Suddenly Blanketing America Man Trapped in Dystopian Nightmare Thanks to AI Surveillance Cameras Flagging His Every Move John Oliver Just Took the AI Industry Behind a Shed and Beat It With a Pipe Wrench OpenAI Hit With Barrage of Lawsuits Over Failure to Report School Shooter Before Massacre Police Are Using AI Camera Networks to Stalk Women Sam Altman Caught in What May Be His Most Spectacular Lie Yet OpenAI in Shambles as IPO Looms A Tiny Town Is Building So Many Data Centers That There’ll Be Almost Nothing Else Left Weird Things Happen When You Give AI Agents Money and Let Them Spend It Sam Altman Issues Grim Apology Top Medical Journal Publishes Searing Article Warning Against Medical AI New Browser Plugin Adds Typos to Your AI-Generated Emails to Make Them Look Real Experts Warn of AI Swarms Hijacking Democracy With Fake Citizens Devious New AI Tool “Clones” Software So That the Original Creator Doesn’t Hold a Copyright Over the New Version Prestigious Wall Street Law Firm Humiliated When Its AI Use Is Discovered in Court Unions Attack AI for Menacing Human Jobs Your Former Employer Is Selling Your Slacks and Emails to Train AI Three Years Ago Today, “Avengers” Director Joe Russo Predicted There Would Be a Fully AI-Generated Movie Within Two Years Palantir’s Employees Are in Crisis
Anthropic Says Claude Turned Evil for a Bizarre Reason
Krystle Verm · 2026-05-13 · via Futurism

A glowing red humanoid face with bright white eyes and the Claude symbol on its forehead, set against a dark purple background. The face has a smooth, almost mask-like appearance with subtle facial features.

Illustration by Tag Hartman-Simkins / Futurism. Source: Getty Images

Sign up to see the future, today

Can’t-miss innovations from the bleeding edge of science and tech

In a classic example of the AI industry’s reputational alchemy, Anthropic has often transformed bad behavior by its flagship model Claude into fresh hype.

When it revealed its Mythos Preview model last month, for example, the company declared that the system had “reached a level of coding capability where they can surpass all but the most skilled humans at finding and exploiting software vulnerabilities.” And last year, it conceded that during the testing of its Claude Opus 4 model, the AI ended up blackmailing a human user upon being threatened with shutdown.

The maneuver was obvious to anyone who’s been watching OpenAI CEO Sam Altman’s antics at Anthropic’s chief rival: the more threatening a problem the AI industry can cook up, the more imminently it can sell its own solutions.

Now, for some reason, Anthropic is relitigating the blackmail incident. Specifically, it’s placing the blame for Claude’s evil behavior on an intriguing villain: the internet at large. Or, to put it another way, it says that humanity — all our journalism and speculation and fiction and social media posts about AI that goes bad — went into Claude’s training data and led the bot astray.

“We started by investigating why Claude chose to blackmail,” the company wrote on X-formerly-Twitter. “We believe the original source of the behavior was internet text that portrays AI as evil and interested in self-preservation. Our post-training at the time wasn’t making it worse — but it also wasn’t making it better.”

Of course, the explicit remit of a company like Anthropic is to develop clever tech that avoids that type of behavioral trap — so a critic might ask why can’t the company take just accountability for the model’s supposed danger, rather than simply blaming the sum output of humankind.

More on Mythos: Top Security Experts Alarmed by Power of Anthropic’s New Hacker AI