惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

J
Java Code Geeks
F
Fortinet All Blogs
云风的 BLOG
云风的 BLOG
MyScale Blog
MyScale Blog
D
DataBreaches.Net
Stack Overflow Blog
Stack Overflow Blog
A
About on SuperTechFans
Google DeepMind News
Google DeepMind News
Microsoft Security Blog
Microsoft Security Blog
腾讯CDC
The GitHub Blog
The GitHub Blog
Jina AI
Jina AI
B
Blog RSS Feed
I
InfoQ
N
Netflix TechBlog - Medium
T
The Blog of Author Tim Ferriss
Microsoft Azure Blog
Microsoft Azure Blog
Recent Announcements
Recent Announcements
GbyAI
GbyAI
H
Help Net Security
L
LangChain Blog
M
MIT News - Artificial intelligence
Y
Y Combinator Blog
aimingoo的专栏
aimingoo的专栏

The Decoder

The AI industry's platform trap is starting to look a lot like Microsoft's OpenAI buys Ona to push Codex toward long-running, autonomous coding tasks Jeff Bezos' AI startup Prometheus closes $12 billion round at a $41 billion valuation Free Deezer tool lets users on any streaming service check their playlists for AI music OpenAI vs. Anthropic: A price war over API tokens is brewing Dario Amodei's new essay reads like a Cold War playbook for the AI age Claude Fable 5: Anthropic admits "wrong tradeoff" after invisibly throttling rival AI researchers Google's new open model DiffusionGemma generates text from noise instead of word by word OpenAI's IPO slips as Altman tells staff to expect a public offering "within the next year" Anthropic study shows AI needs hours, not weeks, to build exploits from security patches OpenAI wants its biggest data center yet, and Nvidia would back the bill Claude Fable 5: The first Mythos model is powerful, expensive, and heavily filtered Germany's National Security Council greenights an AI Safety Institute modeled after the UK's AISI Google's NotebookLM now runs its own cloud computer with code execution and agent-based research Anthropic releases Claude Fable 5 and Mythos 5 with major gains in coding and science Google's Gemini 3.5 Live Translate delivers real-time voice translation across 70+ languages SpaceX wants to put data centers in orbit, and Musk says it's no big deal Landmark German ruling declares Google's AI Overviews are Google's own words and makes it liable for false answers Beijing's $295 billion AI buildout would require 80 percent domestic chips, locking out US suppliers Apple Intelligence gets a second shot with help from Google and Nvidia OpenAI now says "entirely automating everything is not the future we want" OpenAI says going public is "a complicated set of tradeoffs" and is unsure about the timing Microsoft Research's Lens proves detailed captions matter more than raw scale for training efficient image generators Intel gets a second life as Google and Nvidia explore it as a TSMC backup for AI chips Most companies are flying blind on AI spending Frontier Radar #3: How agentic AI is turning tokens into a business metric Instagram AI chatbot breach may have affected over to 20,000 accounts, Meta discloses Microsoft tightens rules for conflict zones after investigation into Israel's military use of Azure Moonshot AI targets a $30 billion valuation, more than six times its late-2025 worth Deepseek topped Ramp's trending software vendors in June 2026 as US companies chase cheaper AI
US government forces Anthropic to disable Claude Fable 5 ...
Matthias Bastian · 2026-06-13 · via The Decoder

The US government has directed Anthropic to shut down access to its most powerful AI models, Fable 5 and Mythos 5, worldwide, citing national security concerns. Anthropic is complying but publicly pushing back.

The export control directive bans all access to Fable 5 and Mythos 5 by foreign nationals, whether they're inside or outside the US. Even Anthropic's own foreign employees are affected.

To comply, Anthropic has to cut off access for all customers worldwide. All other Anthropic models remain available, according to the company's statement. Anthropic calls the move a "misunderstanding" and says it's working to restore access as quickly as possible. The company plans to share more details within 24 hours.

Government claims jailbreak risk, Anthropic disagrees

According to Anthropic, the government believes it has found a method to bypass Fable 5's safety measures. The company says it reviewed a demo of the technique and found it identifies only "a small number of previously known, minor vulnerabilities" that other publicly available models could also detect.

The potential jailbreak—so far only described verbally by the government—boils down to asking the model to read a specific codebase and fix software bugs. Anthropic says it reviewed the report behind the directive and concluded that the capabilities shown are "widely available from other models," including OpenAI's GPT-5.5. Security researchers already use these capabilities daily to protect systems.

Anthropic's own cybersecurity marketing comes back to bite it

Before launch, the US government, the UK AI Safety Institute (UK AISI), private third-party organizations, and internal teams tested the model for thousands of hours combined. The safety measures are "substantially more effective than those of any previously deployed model," Anthropic says. Users even complained they were too restrictive.

No tester has found a universal jailbreak, a method that could broadly bypass the model's safety measures and unlock a wide range of cyber capabilities. But Anthropic also says that perfect jailbreak resistance isn't possible for any model provider right now, a fact well-documented given the sheer number of attack vectors LLMs offer. Every safeguard used across the industry is vulnerable to non-universal jailbreaks that can extract some information in specific cases, Anthropic says.

Knowing this, the company pursued a strategy it calls "defense in depth": keep jailbreaks either narrowly scoped or expensive to pull off, combined with broad monitoring to quickly detect and shut down successful attacks. Part of this strategy includes 30-day data retention for customer data, which Anthropic says creates "real costs for us with customers" but enables jailbreak research and mitigation.

Anyone who previously criticized Anthropic for fear-based marketing can see the irony here. The company spent months loudly warning about the cybersecurity risks of Mythos-class models, working hard to show how superior the model is. Now it has to argue that models already on the market have similar capabilities.

Anthropic warns of a dangerous precedent for the entire industry

Anthropic is complying with the order but making its objections clear. "We disagree that the finding of a narrow potential jailbreak should be cause for recalling a commercial model deployed to hundreds of millions of people." If this standard were applied across the industry, it would effectively halt all new model deployments from every frontier model provider, the company says.

In earlier public statements, Anthropic argued that the government should have the power to block unsafe deployments, but through a legal process that is "transparent, fair, clear, and grounded in technical facts." The current action doesn't meet those principles, the company says, hinting that this could become another chapter in the ongoing clash between Anthropic and the US government.

The US government recently issued a new executive order that lets AI developers submit their models for government safety review before release. Anthropic welcomed that approach, but the process apparently wasn't in place yet when the directive came down.

LLMs remain a weak spot in every cybersecurity setup

Jailbreaks and the related problem of prompt injections have been an unsolved security problem since the early days of large language models. No LLM maker is immune. The vulnerability has been known since at least GPT-3 and affects all LLM-based systems. ChatGPT and Claude can still be attacked through prompt injection under certain conditions, even though their makers have added countermeasures.

Even targeted security efforts have fallen short. About a year ago, Anthropic built a specialized defense against manipulation attempts and put it through a public jailbreaking challenge. After five days, over 300,000 messages, and roughly 3,700 collective work hours, the system was completely cracked, including a universal jailbreak.

AI News Without the Hype – Curated by Humans

Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section.

Subscribe now