惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
爱范儿
爱范儿
H
Help Net Security
V
Visual Studio Blog
J
Java Code Geeks
Stack Overflow Blog
Stack Overflow Blog
Microsoft Security Blog
Microsoft Security Blog
Apple Machine Learning Research
Apple Machine Learning Research
MyScale Blog
MyScale Blog
The Cloudflare Blog
Martin Fowler
Martin Fowler
D
Docker
腾讯CDC
F
Fortinet All Blogs
雷峰网
雷峰网
GbyAI
GbyAI
G
Google Developers Blog
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
Recent Announcements
Recent Announcements
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
Blog — PlanetScale
Blog — PlanetScale
Engineering at Meta
Engineering at Meta
博客园 - 聂微东
博客园 - 叶小钗

Digital Transformation

Aussie Treasurer outlines government goals with AI – productivity up, interest rates down PODCAST: AI won’t replace security teams – it will decide whether they keep you, with Securonix’s Ajay Biyani Op-Ed: If we want to lead the world in AI, we must start in primary school classrooms The Industry Speaks, Part 3: AI Appreciation Day 2026 CBA chief economist doesn’t believe the AI bubble is ‘dotcom 2.0’ The Industry Reacts: Can we make AI work in Australia’s national interest? Visa Visa new AI banking assistant to be adopted by financial firms Australians using AI to guide purchasing finds Zip The industry speaks – part 2: AI Appreciation Day 2026 51% of banks boosting productivity with pilot AI Submissions and nominations are now open for the Australian AI Awards 2026 PODCAST: Trust is the new attack surface, with ThreatLocker APAC director of operations Emile Barakat Need to Know: Aussie workers are increasingly exposing customer data to public AI OpenAI’s Altman won’t accept anything under a $1tn IPO valuation Bendigo Bank has over 3k ideas for agentic AI use AI is running two races – revenue growth and stable economy, says Citi CEO ANZ Bank CIO announces new technology strategy ABC’s Anthropic partnership will see AI generate articles Nine, Microsoft partner for AI use in news Op-Ed: What happens when access to intelligence is no longer your decision? Legal trouble: Misleading AI images could lead to millions in fines PODCAST: Why cyber security’s next battle will be won by speed, with Qualys CEO Sumedh Thakar AI costs lead businesses to rethink their AI investment Bendigo Bank wants to have Australia’s first agentic SOC, but will human workers pay the price? 3 in 4 consumers would ditch a company if it suffered a major cyber attack Bankers warn of central market crash thanks to AI boom Artificial intelligence adoption surged in 2024–25, ABS reports PODCAST: Beware AI and influencers, NSW Rural Fire Service hacked, and say goodbye to the Essential Eight! Productivity Commission says the productivity slump can be cured by AI
Tall tales and code: Anthropic launches Claude Fable 5 an...
David Hollingworth · 2026-06-10 · via Digital Transformation

AI giant Anthropic has released new AI models for both general and private use; however, one expert has warned that “the model will cheerfully produce an exploit to break into a hospital network”.

Anthropic launched a pair of new AI models overnight: the “Mythos-class” Claude Fable 5 and the now out of preview Claude Mythos 5.

Fable 5 – which the company called its most powerful generally available model yet – does come with some caveats, however, with Anthropic launching the model with a set of safeguards to protect against misuse, particularly in areas such as cyber security.

You’re out of free articles for this month

To continue reading the rest of this article, please log in.

Anthropic said that queries on certain topics will be routed to Claude Opus 4.8, its “next most capable model”.

“To release the model both safely and quickly, we’ve tuned these safeguards conservatively – they’ll sometimes catch harmless requests, though they trigger, on average, in less than 5 per cent of sessions,” Anthropic said in a 9 June blog post.

“With more capable models arriving in the coming months, we’re working to improve our safeguards and reduce false positives as quickly as we can.”

As for Claude Mythos 5, the model lacks those safeguards in some areas and is thus still only available under the auspices of Project Glasswing, at least initially. Anthropic has collaborated with the United States government on the rollout.

Claude Mythos 5, according to Anthropic, “has the strongest cyber security capabilities of any model in the world”.

“The capabilities of models like Fable 5 and Mythos 5 have the potential to do profound good for the world,” Anthropic said.

“We’ve seen the beginnings of this in Project Glasswing, where the models have helped cyber defenders secure critically important software.”

Anthropic is claiming its new models beat competitors such as GPT 5.5 and Gemini 3.1 Pro across every metric, from agentic coding to cyber security. However, not everyone is taking Anthropic at face value.

Charles Guillemet, chief technology officer at digital security firm Ledger, said that Anthropic’s reassurances that the models are safe cannot be trusted.

“If you’re reassured that Anthropic has only shipped a ‘safe’ version of Mythos, don’t be. Large language models’ safeguards have repeatedly shown don’t survive contact with even the laziest adversary. Ask politely a few times; frame it as your son’s science-fair project. The model will cheerfully produce an exploit to break into a hospital network,” Guillemet said.

“In reality, attackers have had functionally equivalent capability for months. The proof is in the tidal wave of exploitation we’ve seen around the world, and the price of stolen access on dark markets has never been lower.

“The only layer that makes infrastructure and humans truly resistant to the rapid proliferation of cyber vulnerabilities exposed and exploited by AI is security by design, including formal verification, using hardware-based secure enclaves. In spite of this, individuals and organisations remain slow to update their software stacks.”

Similarly, Andrew Rubin, chief executive and founder at cyber security company Illumio, said the introduction of guardrails is not proof that the problem has been solved, but rather “the companies building these models don’t fully trust where the capability leads”.

“Constraints at the interface don’t change the underlying math; they simply shape how people can interact with it. Attackers won’t operate at that layer. They’ll go straight after the capability itself,” Rubin said.

“And as these tools become more broadly available, the speed and scale of attacks will only increase. The real question isn’t whether guardrails exist – it’s whether defenders are prepared to operate at the same speed.”

Cyber DailyWant to see more stories from trusted news sources?
Make Cyber Daily a preferred news source on Google.

David Hollingworth

David Hollingworth has been writing about technology for over 20 years, and has worked for a range of print and online titles in his career. He is enjoying getting to grips with cyber security, especially when it lets him talk about Lego.