惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
爱范儿
爱范儿
H
Help Net Security
V
Visual Studio Blog
J
Java Code Geeks
Stack Overflow Blog
Stack Overflow Blog
Microsoft Security Blog
Microsoft Security Blog
Apple Machine Learning Research
Apple Machine Learning Research
MyScale Blog
MyScale Blog
The Cloudflare Blog
Martin Fowler
Martin Fowler
D
Docker
腾讯CDC
F
Fortinet All Blogs
雷峰网
雷峰网
GbyAI
GbyAI
G
Google Developers Blog
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
Recent Announcements
Recent Announcements
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
Blog — PlanetScale
Blog — PlanetScale
Engineering at Meta
Engineering at Meta
博客园 - 聂微东
博客园 - 叶小钗

Cyber Daily News

Exclusive: Aussie car part importer Strategic Imports allegedly breached by threat actors New South Wales, other states, investigating Instructure/Canvas data breach Australian Cyber Security Centre warns of ClickFix campaign leveraging Australian infrastructure OpenAI partners with PwC to assist CFOs with AI agents Queensland Department of Education confirms students & staff impacted by ShinyHunters data breach ACMA takes action against SpinTel & Yomojo over mobile number fraud violations The Industry Speaks, Part 1: World Password Day 2026 Qualys and Converge tie cyber insurance pricing to real-time security posture Fakeout: Iranian APT caught hiding behind Chaos ransomware activity Exclusive: Australian energy management firm allegedly breached by SafePay APRA warns of cyber and governance risk due to lagging AI risk management Op-Ed: Australia’s next budget must treat cyber resilience as essential infrastructure Real estate giant Cushman & Wakefield confirms cyber incident, Qilin and ShinyHunters claim attack CrowdStrike expands Project QuiltWorks as more partners join AI security coalition Hacked: ALS discloses cyber incident, unauthorised access to IT systems Microsoft the main target of AI phishing attacks, report uncovers Attackers increasingly turning to trusted security tools to compromise Aussie victims Exclusive: Champion Homes confirms customer data compromised in “cyber event” Australia, Japan commit to partnership to meet cyber security challenges & strengthen cyber defences NSW Treasury cyber incident contained, impact no longer ‘significant’ Report: AI-based data incidents on the rise in Australia WA rental scam surge: Tenants targeted with fake $500 discount trap Aussie Information Commissioner launches Privacy Awareness Week 2026 Unregistered branded text messages to be labelled ‘Unverified’ from 1 July US Federal Reserve outlines AI's influence on the finance sector Exclusive: Major Australian jewellery brand confirms cyber incident Australian government establishes new Cyber Incident Review Board Watch this! Komari server monitor tool abused by hackers Act Now! ACSC warns of active exploitation of cPanel & WHM critical vulnerability Exclusive: Kiwi electrical contractor confirms cyber attack
Tall tales and code: Anthropic launches Claude Fable 5 an...
David Hollingworth · 2026-06-10 · via Cyber Daily News

AI giant Anthropic has released new AI models for both general and private use; however, one expert has warned that “the model will cheerfully produce an exploit to break into a hospital network”.

Anthropic launched a pair of new AI models overnight: the “Mythos-class” Claude Fable 5 and the now out of preview Claude Mythos 5.

Fable 5 – which the company called its most powerful generally available model yet – does come with some caveats, however, with Anthropic launching the model with a set of safeguards to protect against misuse, particularly in areas such as cyber security.

You’re out of free articles for this month

To continue reading the rest of this article, please log in.

Anthropic said that queries on certain topics will be routed to Claude Opus 4.8, its “next most capable model”.

“To release the model both safely and quickly, we’ve tuned these safeguards conservatively – they’ll sometimes catch harmless requests, though they trigger, on average, in less than 5 per cent of sessions,” Anthropic said in a 9 June blog post.

“With more capable models arriving in the coming months, we’re working to improve our safeguards and reduce false positives as quickly as we can.”

As for Claude Mythos 5, the model lacks those safeguards in some areas and is thus still only available under the auspices of Project Glasswing, at least initially. Anthropic has collaborated with the United States government on the rollout.

Claude Mythos 5, according to Anthropic, “has the strongest cyber security capabilities of any model in the world”.

“The capabilities of models like Fable 5 and Mythos 5 have the potential to do profound good for the world,” Anthropic said.

“We’ve seen the beginnings of this in Project Glasswing, where the models have helped cyber defenders secure critically important software.”

Anthropic is claiming its new models beat competitors such as GPT 5.5 and Gemini 3.1 Pro across every metric, from agentic coding to cyber security. However, not everyone is taking Anthropic at face value.

Charles Guillemet, chief technology officer at digital security firm Ledger, said that Anthropic’s reassurances that the models are safe cannot be trusted.

“If you’re reassured that Anthropic has only shipped a ‘safe’ version of Mythos, don’t be. Large language models’ safeguards have repeatedly shown don’t survive contact with even the laziest adversary. Ask politely a few times; frame it as your son’s science-fair project. The model will cheerfully produce an exploit to break into a hospital network,” Guillemet said.

“In reality, attackers have had functionally equivalent capability for months. The proof is in the tidal wave of exploitation we’ve seen around the world, and the price of stolen access on dark markets has never been lower.

“The only layer that makes infrastructure and humans truly resistant to the rapid proliferation of cyber vulnerabilities exposed and exploited by AI is security by design, including formal verification, using hardware-based secure enclaves. In spite of this, individuals and organisations remain slow to update their software stacks.”

Similarly, Andrew Rubin, chief executive and founder at cyber security company Illumio, said the introduction of guardrails is not proof that the problem has been solved, but rather “the companies building these models don’t fully trust where the capability leads”.

“Constraints at the interface don’t change the underlying math; they simply shape how people can interact with it. Attackers won’t operate at that layer. They’ll go straight after the capability itself,” Rubin said.

“And as these tools become more broadly available, the speed and scale of attacks will only increase. The real question isn’t whether guardrails exist – it’s whether defenders are prepared to operate at the same speed.”

Cyber DailyWant to see more stories from trusted news sources?
Make Cyber Daily a preferred news source on Google.

David Hollingworth

David Hollingworth has been writing about technology for over 20 years, and has worked for a range of print and online titles in his career. He is enjoying getting to grips with cyber security, especially when it lets him talk about Lego.