惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

J
Java Code Geeks
Jina AI
Jina AI
小众软件
小众软件
WordPress大学
WordPress大学
Last Week in AI
Last Week in AI
美团技术团队
V
V2EX
酷 壳 – CoolShell
酷 壳 – CoolShell
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
博客园 - 聂微东
博客园 - 【当耐特】
人人都是产品经理
人人都是产品经理
雷峰网
雷峰网
博客园 - 司徒正美
量子位
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
宝玉的分享
宝玉的分享
月光博客
月光博客
IT之家
IT之家
博客园 - 三生石上(FineUI控件)
大猫的无限游戏
大猫的无限游戏
T
Tailwind CSS Blog
博客园 - Franky

Digital Transformation

The Industry Reacts: Can we make AI work in Australia’s national interest? Visa Visa new AI banking assistant to be adopted by financial firms Australians using AI to guide purchasing finds Zip The industry speaks – part 2: AI Appreciation Day 2026 51% of banks boosting productivity with pilot AI Submissions and nominations are now open for the Australian AI Awards 2026 PODCAST: Trust is the new attack surface, with ThreatLocker APAC director of operations Emile Barakat Need to Know: Aussie workers are increasingly exposing customer data to public AI OpenAI’s Altman won’t accept anything under a $1tn IPO valuation Bendigo Bank has over 3k ideas for agentic AI use AI is running two races – revenue growth and stable economy, says Citi CEO ANZ Bank CIO announces new technology strategy ABC’s Anthropic partnership will see AI generate articles Nine, Microsoft partner for AI use in news Op-Ed: What happens when access to intelligence is no longer your decision? Legal trouble: Misleading AI images could lead to millions in fines PODCAST: Why cyber security’s next battle will be won by speed, with Qualys CEO Sumedh Thakar AI costs lead businesses to rethink their AI investment Bendigo Bank wants to have Australia’s first agentic SOC, but will human workers pay the price? 3 in 4 consumers would ditch a company if it suffered a major cyber attack Bankers warn of central market crash thanks to AI boom Artificial intelligence adoption surged in 2024–25, ABS reports PODCAST: Beware AI and influencers, NSW Rural Fire Service hacked, and say goodbye to the Essential Eight! Productivity Commission says the productivity slump can be cured by AI APRA instructs major lenders to share cyber insights AI law firm wins court case in legal profession first Major UK bank to launch over 1,000 new AI roles CPA warns against using AI tools and ‘finfluencer’ advice when tax time comes Suncorp trials agentic AI in insurance claims processing
Tall tales and code: Anthropic launches Claude Fable 5 an...
David Hollingworth · 2026-06-10 · via Digital Transformation

AI giant Anthropic has released new AI models for both general and private use; however, one expert has warned that “the model will cheerfully produce an exploit to break into a hospital network”.

Anthropic launched a pair of new AI models overnight: the “Mythos-class” Claude Fable 5 and the now out of preview Claude Mythos 5.

Fable 5 – which the company called its most powerful generally available model yet – does come with some caveats, however, with Anthropic launching the model with a set of safeguards to protect against misuse, particularly in areas such as cyber security.

You’re out of free articles for this month

To continue reading the rest of this article, please log in.

Anthropic said that queries on certain topics will be routed to Claude Opus 4.8, its “next most capable model”.

“To release the model both safely and quickly, we’ve tuned these safeguards conservatively – they’ll sometimes catch harmless requests, though they trigger, on average, in less than 5 per cent of sessions,” Anthropic said in a 9 June blog post.

“With more capable models arriving in the coming months, we’re working to improve our safeguards and reduce false positives as quickly as we can.”

As for Claude Mythos 5, the model lacks those safeguards in some areas and is thus still only available under the auspices of Project Glasswing, at least initially. Anthropic has collaborated with the United States government on the rollout.

Claude Mythos 5, according to Anthropic, “has the strongest cyber security capabilities of any model in the world”.

“The capabilities of models like Fable 5 and Mythos 5 have the potential to do profound good for the world,” Anthropic said.

“We’ve seen the beginnings of this in Project Glasswing, where the models have helped cyber defenders secure critically important software.”

Anthropic is claiming its new models beat competitors such as GPT 5.5 and Gemini 3.1 Pro across every metric, from agentic coding to cyber security. However, not everyone is taking Anthropic at face value.

Charles Guillemet, chief technology officer at digital security firm Ledger, said that Anthropic’s reassurances that the models are safe cannot be trusted.

“If you’re reassured that Anthropic has only shipped a ‘safe’ version of Mythos, don’t be. Large language models’ safeguards have repeatedly shown don’t survive contact with even the laziest adversary. Ask politely a few times; frame it as your son’s science-fair project. The model will cheerfully produce an exploit to break into a hospital network,” Guillemet said.

“In reality, attackers have had functionally equivalent capability for months. The proof is in the tidal wave of exploitation we’ve seen around the world, and the price of stolen access on dark markets has never been lower.

“The only layer that makes infrastructure and humans truly resistant to the rapid proliferation of cyber vulnerabilities exposed and exploited by AI is security by design, including formal verification, using hardware-based secure enclaves. In spite of this, individuals and organisations remain slow to update their software stacks.”

Similarly, Andrew Rubin, chief executive and founder at cyber security company Illumio, said the introduction of guardrails is not proof that the problem has been solved, but rather “the companies building these models don’t fully trust where the capability leads”.

“Constraints at the interface don’t change the underlying math; they simply shape how people can interact with it. Attackers won’t operate at that layer. They’ll go straight after the capability itself,” Rubin said.

“And as these tools become more broadly available, the speed and scale of attacks will only increase. The real question isn’t whether guardrails exist – it’s whether defenders are prepared to operate at the same speed.”

Cyber DailyWant to see more stories from trusted news sources?
Make Cyber Daily a preferred news source on Google.

David Hollingworth

David Hollingworth has been writing about technology for over 20 years, and has worked for a range of print and online titles in his career. He is enjoying getting to grips with cyber security, especially when it lets him talk about Lego.