惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Recent Announcements
Recent Announcements
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
MongoDB | Blog
MongoDB | Blog
H
Help Net Security
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
人人都是产品经理
人人都是产品经理
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
The GitHub Blog
The GitHub Blog
V
V2EX
Microsoft Security Blog
Microsoft Security Blog
V
Visual Studio Blog
A
About on SuperTechFans
博客园_首页
L
LangChain Blog
量子位
雷峰网
雷峰网
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
Jina AI
Jina AI
月光博客
月光博客
阮一峰的网络日志
阮一峰的网络日志
博客园 - 聂微东
Microsoft Azure Blog
Microsoft Azure Blog
M
MIT News - Artificial intelligence
N
Netflix TechBlog - Medium

Futurism

Meta Installing Software on Employee Computers to Track Everything They Do, Feed the Data to AI Concern Grows That AI Is Damaging Users’ Cognitive Abilities JPMorganChase Data Center Gets $77 Million Handout to Create Grand Total of One Job Nvidia CEO Loses His Cool at Tough Question CEO of $1.5 Billion AI Startup Accused of Massive Fraud by Justice Department Palantir Issues Ominous Corporate Manifesto Madison Square Garden Reportedly Used Facial Recognition to Stalk Trans Woman For Two Years The Florida Mass Shooter’s Conversations With ChatGPT Are Worse Than You Could Possibly Imagine China Is Starting to Pull Ahead of US in AI Race AI Company Known for Teen Suicides Launches New Feature to Turn Books Into Roleplaying Experiences Study Finds AI Use Eats Away at Users’ Confidence in Their Own Brains Democrats Warned Not to Upset Multi-Million Dollar AI Lobbyists, Even Though It’d Be a Slam Dunk With Voters City Council Wrecked in Voter Bloodbath After Allowing New Data Center Mother Reportedly Doesn’t Know Her Son Died Because She’s Been Talking to an AI Version of Him Things You Told ChatGPT or Claude My Have Already Doomed You in Court Millions of Americans Are Talking to AI Instead of Going to the Doctor, and It’s Giving Them Horrendously Flawed Medical Advice There Are Signs of a Massive AI Backlash A Prominent PR Firm Is Running a Fake News Site That’s Plagiarizing Original Journalism at Incredible Scale Fury Erupts as Val Kilmer’s Estate Announces Starring Role in AI Film Made From Beyond the Grave Allbirds Stock Now Crashing as Reality Sets in About Its Delusional AI Pivot NAACP Sues Elon Over His Noxious AI Data Center Top Security Experts Alarmed by Power of Anthropic’s New Hacker AI Teens Alarmed at What AI Is Doing to Their Minds What It Really Means That a Failing Shoe Brand “Pivoted to AI” and Its Stock Soared 700 Percent Starbucks’ Baffling ChatGPT Collab Treats Customers Like Empty, Soulless Venti Cups ChatGPT’s “Honest Reaction” to a “Song” Composed Entirely of Gas-Passing Noises Will Make You Question Whether It’s Honestly Evaluating Your Other Brilliant Ideas AI Is Turning Workplaces Into Hopeless Gridlock Companies Just Learned a Brutal Lesson About Training AI to Do Human Jobs Berklee College of Music Students Furious That It’s Offering an AI “Songwriting” Class Usually, Young People Embrace New Technology. Gen Z’s Attitude Toward AI Should Worry the Entire Tech Industry
Anthropic Says Claude Turned Evil for a Bizarre Reason
Krystle Verm · 2026-05-13 · via Futurism

A glowing red humanoid face with bright white eyes and the Claude symbol on its forehead, set against a dark purple background. The face has a smooth, almost mask-like appearance with subtle facial features.

Illustration by Tag Hartman-Simkins / Futurism. Source: Getty Images

Sign up to see the future, today

Can’t-miss innovations from the bleeding edge of science and tech

In a classic example of the AI industry’s reputational alchemy, Anthropic has often transformed bad behavior by its flagship model Claude into fresh hype.

When it revealed its Mythos Preview model last month, for example, the company declared that the system had “reached a level of coding capability where they can surpass all but the most skilled humans at finding and exploiting software vulnerabilities.” And last year, it conceded that during the testing of its Claude Opus 4 model, the AI ended up blackmailing a human user upon being threatened with shutdown.

The maneuver was obvious to anyone who’s been watching OpenAI CEO Sam Altman’s antics at Anthropic’s chief rival: the more threatening a problem the AI industry can cook up, the more imminently it can sell its own solutions.

Now, for some reason, Anthropic is relitigating the blackmail incident. Specifically, it’s placing the blame for Claude’s evil behavior on an intriguing villain: the internet at large. Or, to put it another way, it says that humanity — all our journalism and speculation and fiction and social media posts about AI that goes bad — went into Claude’s training data and led the bot astray.

“We started by investigating why Claude chose to blackmail,” the company wrote on X-formerly-Twitter. “We believe the original source of the behavior was internet text that portrays AI as evil and interested in self-preservation. Our post-training at the time wasn’t making it worse — but it also wasn’t making it better.”

Of course, the explicit remit of a company like Anthropic is to develop clever tech that avoids that type of behavioral trap — so a critic might ask why can’t the company take just accountability for the model’s supposed danger, rather than simply blaming the sum output of humankind.

More on Mythos: Top Security Experts Alarmed by Power of Anthropic’s New Hacker AI