惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

S
SegmentFault 最新的问题
博客园 - 三生石上(FineUI控件)
WordPress大学
WordPress大学
博客园 - 【当耐特】
月光博客
月光博客
Vercel News
Vercel News
D
Docker
I
InfoQ
Apple Machine Learning Research
Apple Machine Learning Research
博客园 - 叶小钗
MongoDB | Blog
MongoDB | Blog
GbyAI
GbyAI
有赞技术团队
有赞技术团队
雷峰网
雷峰网
博客园 - 聂微东
小众软件
小众软件
Y
Y Combinator Blog
腾讯CDC
L
LangChain Blog
The GitHub Blog
The GitHub Blog
宝玉的分享
宝玉的分享
Stack Overflow Blog
Stack Overflow Blog
大猫的无限游戏
大猫的无限游戏
T
The Blog of Author Tim Ferriss

Gadget Review

Bernie Sanders Wants You to Own Half of OpenAI - And He's Not Kidding - Gadget Review California Bill Strikes Back Against Disappearing Video Games - Gadget Review Japan Cracks 6G's Speed Barrier With 112 Gbps Wireless Breakthrough - Gadget Review 31 Amazon Kitchen Tools and Gadgets That Make Prep Time a Breeze The 559-Mile Mic-Drop: Why BMW’s New i3 Just Made Tesla’s Range Look Like A Toy - Gadget Review 20 Genius Camping Gadgets That Will Help Make Summer Camping Easier Tesla Patents Transform Glass Roofs Into Smart Air Conditioners - Gadget Review Is Anthropic’s “Benefit Corp” Structure An Investor’s Worst Nightmare? - Gadget Review How An AI Weather Startup Just Beat the World’s Greatest Supercomputers - Gadget Review Dell's New XPS 13 Is Directly Targeting The MacBook Neo - Gadget Review 13 Smart Home Gadgets for True Local Control (No Cloud Needed!) Florida Sues OpenAI and CEO Sam Altman - Why Florida Is Treating AI Chatbots as "Hazardous Products" - Gadget Review Malaysia’s Scorched-Earth Policy Against Under-16 Social Media Access - Ban Carries Fines Up To $2.5 Million - Gadget Review PlayStation's Wireless Fight Stick and Latest Gaming Monitor Hits This August - Gadget Review How Meta's Chatbot Handed Over Million-Dollar Instagram Accounts To Attackers - Gadget Review 11 Home Security Gadgets That Help Safeguard Your Sanctuary Engineer Builds AI-Powered Laser System That Targets & Hunts Mosquitoes at Home - Gadget Review DuckDuckGo's No-AI Search Extensions Surge as Users Flee Google's AI Overhaul - Gadget Review Google Wants to Release 32 Million "Infected" Mosquitoes Into The Wild - Gadget Review Nvidia Is Bringing AI Power To Your Desk With New Superchip - Gadget Review Tech CEOs Are Using AI as the Perfect Scapegoat for Mass Layoffs - Gadget Review Wix Cuts 1,000 Jobs, Citing AI Evolution and Currency Pressures - Gadget Review Teen's Bluetooth Speaker Named "BOMB" Forces Flight U-Turn Mid-Atlantic - Gadget Review China's Humanoid Robots Sort 1,200 Postal Packages Per Hour - Gadget Review California Senate Passes Historic Ban on AI Chatbot Toys - Gadget Review Professor Declares War on AI: Will Fail Any Student Who Uses It - Gadget Review UK Military Looks At Allowing Lethal Strikes With Zero Human Intervention - Gadget Review Chinese EVs Are Tanking in Value - Gadget Review Japanese Researchers Create Chip That Could Run 1,000x Faster, Near-Zero Heat - Gadget Review Total Immobility: Why A Single Targeted Cyberattack Could Leave Every EV In Your City Stranded - Gadget Review
The White House Wants Anthropic to Block All Jailbreaks. ...
Nikshep Myle · 2026-06-18 · via Gadget Review

Commerce Department forced Claude Fable 5 offline June 12 under export rules, then demanded a jailbreak fix security researchers call technically impossible

On June 12, the Commerce Department did something it had never done before: it used Export Administration Regulations to force a commercial AI model offline. Not a chip. Not manufacturing equipment. A piece of software. Anthropic’s Claude Fable 5 — a consumer-facing assistant built on top of the more powerful Mythos 5 system — was pulled from global access after officials concluded its guardrails could be bypassed to expose Mythos’s cybersecurity reasoning capabilities, according to the Cloud Security Alliance. Because Anthropic couldn’t reliably distinguish foreign nationals from U.S. users in real time, it shut down both models for everyone. Every frontier AI lab just got put on notice.

What the Government Actually Wants

The NSA says jailbreaks exist. Anthropic says the fix being demanded would freeze the entire industry.

The White House position, as reported by the Washington Examiner, is blunt: Anthropic must proactively test its own models, patch jailbreaks, and flag findings to the government. Officials say they can’t staff a permanent jailbreak-hunting operation across every commercial AI product. In practice, this amounts to demanding zero exploitable gaps — the price of getting Fable 5 back online.

Here’s what the “jailbreak” actually looked like. According to Simon Willison’s independent analysis, researchers asked the model to fix vulnerable code, then manually converted those fixes into exploit-testing scripts. Fable refused the direct security-review request but complied with the reframed “fix this” prompt — because that’s what good coding assistants do.

Anthropic says it received only verbal evidence of a narrow, non-universal technique. The company stated, per Fortune, that “if this standard were applied across the industry, we believe it would essentially halt all new model deployments for all frontier model providers.”

Why Experts Say the Demand Doesn’t Square With Reality

Asking a model to forget what it knows is like asking Google Maps to forget where the roads are.

Large language models are trained to reason flexibly and follow instructions. Guardrails are linguistic restrictions layered on top of knowledge that still exists inside the system. Sufficiently clever prompts — or future AI systems deployed as attackers — can search prompt-space faster than any human red team. Removing vulnerability-discovery capabilities wholesale would gut the exact defensive security work that makes these tools valuable.

Guardrails aren’t a lock. They’re a screen door.

Dozens of cybersecurity leaders have sounded alarms, as reported by Axios. State-backed offensive teams will access equivalent capabilities regardless — through alternative models, covert channels, or homegrown systems. Pulling Fable 5 creates a lopsided dynamic: defenders get hobbled while well-resourced attackers adapt without missing a beat. This pattern echoes broader tech scandals in which industry accountability lagged far behind government intervention.

What Comes Next

The precedent is set, and the realistic path forward looks nothing like what Washington is currently demanding.

Any model exhibiting strong cyber, bio, or chem capabilities could now be reclassified and pulled overnight — your risk calculus for building on frontier AI just shifted. Moves like this parallel how Europe restricts major cloud providers from handling sensitive government data, signaling a global tightening of the leash on frontier AI. The realistic answer isn’t “block all jailbreaks.” It’s managed risk through:

  • access controls
  • rigorous logging
  • monitored workflows

Governments demanding perfect control of imperfect systems don’t get safety. They get paralysis.