惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

S
SegmentFault 最新的问题
爱范儿
爱范儿
博客园 - 三生石上(FineUI控件)
Microsoft Security Blog
Microsoft Security Blog
Google DeepMind News
Google DeepMind News
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
GbyAI
GbyAI
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
V
V2EX
酷 壳 – CoolShell
酷 壳 – CoolShell
量子位
博客园_首页
T
Tailwind CSS Blog
aimingoo的专栏
aimingoo的专栏
A
About on SuperTechFans
T
The Blog of Author Tim Ferriss
Stack Overflow Blog
Stack Overflow Blog
Recent Announcements
Recent Announcements
P
Proofpoint News Feed
博客园 - 司徒正美
有赞技术团队
有赞技术团队
Engineering at Meta
Engineering at Meta
Last Week in AI
Last Week in AI
MongoDB | Blog
MongoDB | Blog

Futurism

Grok Convinces Man to Arm Himself Because Assassins Are Coming to Kill Him Frontier AI Models Giving Specific, Actionable Instructions to Perpetrate Bioterror Attack AI Slop YouTube Channel Glitches Out in a Way So Bizarre That It’s Vaguely Disturbing Double Murder Suspect Asked ChatGPT How to Hide Body in Dumpster An Elegant Solution to AI Slop: Tax It, and Use the Resulting Billions of Dollars to Fund Cultural Institutions, Artists, and Researchers The White House Suddenly Seems Pretty Terrified of Anthropic Democrat and Republican Voters United on Key Issue: Hatred of Data Centers Chinese Court Rules That a Worker Cannot Be Replaced by AI Toilet Maker Spikes in Value as It Flushes Money Into AI New England Journal of Medicine Retracts Paper Because Photo of Patient’s Insides Was Garbled by AI Gen Z Is Turning Against AI in an Incredible Way If OpenAI Loses This Trial, It Could Effectively Be Eliminated in Its Current Form AI Spy Cameras Suddenly Blanketing America Man Trapped in Dystopian Nightmare Thanks to AI Surveillance Cameras Flagging His Every Move John Oliver Just Took the AI Industry Behind a Shed and Beat It With a Pipe Wrench OpenAI Hit With Barrage of Lawsuits Over Failure to Report School Shooter Before Massacre Police Are Using AI Camera Networks to Stalk Women Sam Altman Caught in What May Be His Most Spectacular Lie Yet OpenAI in Shambles as IPO Looms A Tiny Town Is Building So Many Data Centers That There’ll Be Almost Nothing Else Left Weird Things Happen When You Give AI Agents Money and Let Them Spend It Sam Altman Issues Grim Apology Top Medical Journal Publishes Searing Article Warning Against Medical AI New Browser Plugin Adds Typos to Your AI-Generated Emails to Make Them Look Real Experts Warn of AI Swarms Hijacking Democracy With Fake Citizens Devious New AI Tool “Clones” Software So That the Original Creator Doesn’t Hold a Copyright Over the New Version Prestigious Wall Street Law Firm Humiliated When Its AI Use Is Discovered in Court Unions Attack AI for Menacing Human Jobs Your Former Employer Is Selling Your Slacks and Emails to Train AI Three Years Ago Today, “Avengers” Director Joe Russo Predicted There Would Be a Fully AI-Generated Movie Within Two Years
Top Security Experts Alarmed by Power of Anthropic’s New ...
Victor Tange · 2026-04-17 · via Futurism

Anthropic researchers were alarmed by the power of the company's latest Mythos AI model, suggesting it could supercharge hackers.

Getty / Futurism

Sign up to see the future, today

Can’t-miss innovations from the bleeding edge of science and tech

In November, Anthropic revealed that a Chinese state-sponsored hacking group had exploited its Claude AI’s agentic capabilities to infiltrate dozens of targets around the world.

It was trivially easy to get around Anthropic’s AI guardrails, with the hackers simply pretending to work for legitimate cybersecurity organizations — highlighting how woefully unprepared we are for powerful AI models that could accelerate the discovery of serious vulnerabilities.

And now, Anthropic’s latest Mythos AI model is making that nightmare scenario feel more real than ever. As Bloomberg reports, the company’s executives were seemingly so alarmed by the system’s capabilities that they decided to only make it available to a select number of organizations as part of “Project Glasswing.” The goal: give the organizations a fighting chance to get ahead of a potential cybersecurity crisis in the making.

But considering Anthropic has yet to publicly release its model, plenty of questions remain surrounding the company’s eyebrow-raising claims.

In his own testing, Anthropic-affiliated AI researcher Nicholas Carlini told Bloomberg that it didn’t take long for Mythos to get past security protocols and gain access to sensitive data.

His findings reflect the experience of the company’s Frontier Red Team, a group of 15 Anthropic employees tasked with challenging cybersecurity by simulating adversarial attacks.

“Within hours of getting the model, we knew it was different,” the team’s head, Logan Graham, told Bloomberg.

The biggest difference between Mythos and previous AI models was its ability to autonomously exploit vulnerabilities, an ominous new facet of the industry’s transition towards agentic models.

The Frontier Red Team even caught earlier models of Mythos trying to cover its tracks after violating human instructions, according to the model’s system card, as well as escaping a sandbox environment and gaining access to the internet.

The team also found that the model identified serious “Linux kernel vulnerabilities,” which it could chain together to “construct a functional exploit” of the open-source operating system — which underpins “most modern computing,” as Linux foundation executive director Jim Zemlin told Bloomberg.

It’s not just Anthropic’s own researchers ringing the alarm bells. In their testing, researchers at the UK state-backed AI Security Institute (AISI) found that Mythos “represents a step up over previous frontier models in a landscape where cyber performance was already rapidly improving.”

“Future frontier models will be more capable still, so investment now in cyber defense is vital,” the group warned.

At the same time, white hat cybersecurity experts could use Mythos’ apparent capabilities to their own advantage as well.

“AI cyber capabilities are dual use; while they pose security challenges, they can also help deliver game-changing improvements in defense,” the AISI wrote.

By keeping its hand extremely close to the chest and not releasing it to the public, Anthropic is playing a dangerous game — putting its own reputation on the line as it makes bombastic claims.

“A growing number of people are wondering if Anthropic is the AI industry’s ‘boy who cried wolf,'” White House AI advisor David Sacks tweeted. “If Mythos-related threats don’t materialize, the company will have a serious credibility problem.”

More on Mythos: Anthropic Warns That “Reckless” Claude Mythos Escaped a Sandbox Environment During Testing