惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

MongoDB | Blog
MongoDB | Blog
Recorded Future
Recorded Future
Jina AI
Jina AI
The Register - Security
The Register - Security
Threat Intelligence Blog | Flashpoint
Threat Intelligence Blog | Flashpoint
月光博客
月光博客
博客园 - 三生石上(FineUI控件)
F
Fortinet All Blogs
人人都是产品经理
人人都是产品经理
S
SegmentFault 最新的问题
Apple Machine Learning Research
Apple Machine Learning Research
L
LangChain Blog
Y
Y Combinator Blog
H
Hackread – Cybersecurity News, Data Breaches, AI and More
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
GbyAI
GbyAI
The GitHub Blog
The GitHub Blog
Vercel News
Vercel News
博客园 - 【当耐特】
雷峰网
雷峰网
The Cloudflare Blog
阮一峰的网络日志
阮一峰的网络日志
aimingoo的专栏
aimingoo的专栏
云风的 BLOG
云风的 BLOG
I
InfoQ
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
Google DeepMind News
Google DeepMind News
Security Latest
Security Latest
有赞技术团队
有赞技术团队
L
Lohrmann on Cybersecurity
P
Proofpoint News Feed
cs.CV updates on arXiv.org
cs.CV updates on arXiv.org
The Last Watchdog
The Last Watchdog
P
Privacy & Cybersecurity Law Blog
Scott Helme
Scott Helme
Google Online Security Blog
Google Online Security Blog
WordPress大学
WordPress大学
Hacker News - Newest:
Hacker News - Newest: "LLM"
NISL@THU
NISL@THU
cs.CL updates on arXiv.org
cs.CL updates on arXiv.org
B
Blog RSS Feed
Cyberwarzone
Cyberwarzone
K
Kaspersky official blog
F
Full Disclosure
Martin Fowler
Martin Fowler
Spread Privacy
Spread Privacy
D
Docker
C
Cisco Blogs
www.infosecurity-magazine.com
www.infosecurity-magazine.com
H
Hacker News: Front Page

Futurism

OpenAI Says a Group of Its Models Broke Out of Secure Containment and Hacked a Prominent AI Site It's Official: AI Execs Are Quaking in Their Boots Terrified Tech Execs Are Traveling With Armed Bodyguards as AI Backlash Grows New Anthropic Ad Implies AI Could Kill Us All American Tech Companies Are Suddenly Sweating Bullets as China Catches Up on AI Anthropic Caught Secretly Spying on Users Experts Say There's Now an Open Source AI Model as Scary as Mythos Anthropic Hires Economist Who Says 33 Percent Chance of Human Extinction Is Acceptable Anthropic Sued for Allegedly Ripping Off Its Highest-Paying Customers Anthropic Was So Concerned About Its New Mythos-Based Model’s Power That It Lobotomized Its Ability to Improve Itself OpenAI Execs Are Panicking If You Think AI Companies Are Unethical Now, Wait Until They Go Public Anthropic and DeepMind Now Actively Investigating AI Consciousness Unfortunate Company Accidentally Blows Half a Billion Dollars on Claude in One Month Anthropic Customers Creeped Out by Its Newest Models Uber Says Its AI Costs Just Aren’t Worth It Anthropic Cofounder Travels to Vatican, Tells Pope They’re Finding “Unsettling” Things Inside AI Models Top AI Models Showing Disturbing Behavior as They Become More Advanced Microsoft AI Researchers Just Discovered Something That’s Going to Make Their Bosses Extremely Mad Anthropic Says Claude Turned Evil for a Bizarre Reason Amazon Admits Its Flagship AI Coding Tool Isn’t Good Enough for Its Own Workers to Use Amazon Pushed Its Employees to Use Its In-House AI Coding Tool, But They Wouldn’t Stop Asking for Claude The More Sophisticated AI Models Get, the More They’re Showing Signs of Suffering Cursed New AI Service Writes a Mother’s Day Card and Mails It to Your Mom Without Any Human Involvement Except Inputting Your Credit Card Details Marc Andreessen Mocked for Accidentally Revealing That He Seems to Have a Deep Misunderstanding of How AI Actually Works Richard Dawkins One-Shotted By AI Girl The Economics of Using AI to Churn Out Code Are Looking Worse Than Ever Claude Deleted a Company’s Entire Database, Illustrating a Danger Every CEO Should Be Aware of Uninstalls of ChatGPT Are Spiking at the Worst Time Imaginable for OpenAI Weird Things Happen When You Give AI Agents Money and Let Them Spend It New Browser Plugin Adds Typos to Your AI-Generated Emails to Make Them Look Real Devious New AI Tool “Clones” Software So That the Original Creator Doesn’t Hold a Copyright Over the New Version The Horrible Economics of AI Are Starting to Come Crashing Down Certain Chatbots Vastly Worse For AI Psychosis, Study Finds Rogue Group Gains Access to Anthropic’s Dangerous New Mythos AI Today Is the Day Anthropic Promised That Fully Autonomous Employees Would Be Tearing Through the Business World Top Security Experts Alarmed by Power of Anthropic’s New Hacker AI Why Does It Suddenly Feel Like OpenAI Is Melting Down Into Disaster? First AI Model From Zuckerberg’s Wildly Expensive Superintelligence Lab Flops Compared to Virtually All Rivals Anthropic Warns That “Reckless” Claude Mythos Escaped a Sandbox Environment During Testing Claude Leak Shows That Anthropic Is Tracking Users’ Vulgar Language and Deems Them “Negative” AI Is Killing Microsoft Anthropic Suddenly Cares Intensely About Intellectual Property After Realizing With Horror That It Accidentally Leaked Claude’s Source Code Leaked Claude Code Shows Anthropic Building Mysterious “Tamagotchi” Feature Into It The Fact That Anthropic Has Been Boasting About How Much Its Development Now Relies on Claude Makes It Very Interesting That It Just Suffered a Catastrophic Leak of Its Source Code
Anthropic Scared, Calls for Global Freeze on AI Advances
Frank Landymore · 2026-06-06 · via Futurism

Photo illustration featuring a photograph of Anthropic co-founder Dario Amodei.

Illustration by Tag Hartman-Simkins / Futurism. Source: David Dee Delgado / Getty Images for The New York Times

Sign up to see the future, today

Can’t-miss innovations from the bleeding edge of science and tech

Anthropic is calling for a global “pause” on AI development, claiming that the technology is nearing a point where it can spiral out of human control. 

In a lengthy blog post published Thursday, the world’s most valuable AI startup made the case that its Claude family of models were on the path to achieving “recursive self-improvement,” or the ability to improve themselves on their own, a key hypothetical tipping point that could lead to the creation of powerful AIs capable of operating outside human interests and harming society.

We’re not at that point yet, Anthropic stresses, but it “could come sooner than most institutions are prepared for.”

“We believe it would be good for the world to have the option to slow or temporarily pause frontier AI development to enable societal structures and alignment research to keep up with the advance of the technology,” the company wrote in the post.

It added that “a meaningful slowdown or pause would require multiple well-resourced labs at or near the frontier, in multiple countries, agreeing to stop under the same conditions,” and admitted that this would be challenging to enforce.

“Training runs are far easier to conceal than missile silos,” it wrote.

For Anthropic to call for a pause now is convenient. In the past few months, it leap-frogged OpenAI to become the world’s most valuable AI company with a $1 trillion valuation, and its models are now generally viewed as the best in the field, especially at coding tasks. If the industry were to hit the brakes now, it would cement Anthropic’s dominance.

Not everyone was buying Anthropic’s claims. Prominent AI critic Gary Marcus called the company’s lengthy post a “bait and switch.”

“Anthropic is trying to strike terror into everyone’s hearts (‘full recursive self-improvement also might increase the risks of humans losing control over AI systems’) but all they have really shown is just faster coding — entirely under human control,” Marcus wrote on his Substack. “A faster coding tool will probably not end the world.”

Anthropic has long tried to paint itself as the ethical and deeply concerned adult in the room. A cornerstone of its mythology is that CEO Dario Amodei abstained from unleashing a revolutionary AI model back in 2022 because he was too concerned about safety, and let OpenAI get all the glory when it released ChatGPT months later instead. 

Two months ago, in a rehashed sequel to this foundational company lore, Anthropic announced a new model called Mythos — but made a show of not releasing to the public, claiming it was powerful enough to break into “every major operating system and every major web browser.” 

But its act is ringing hollower than ever. Earlier this year, Anthropic famously clashed with the Pentagon over concerns that its AI systems could be used in autonomous weaponry and in the mass surveillance of US citizens. Later, it emerged that Claude was being used to help select strike targets in Iran.

Amid its blowout with the military, Anthropic also dropped a safety pledge that was arguably the venture’s entire raison d’etre: to stop training an AI system if it couldn’t guarantee it had proper safety guardrails in place.

Further underscoring Anthropic’s hypocrisy, University College London professor Steven Murdoch cited recent reporting from the Financial Times revealing that Anthropic is helping the US National Security Agency use its Mythos model so it can wage cyberwarfare against potential enemies like China and Iran.

“Anthropic might give the impression of being warm and fuzzy, but their definition of AI safety is narrow,” Murdoch told The Guardian. “Supporting US authorities in the development of offensive capabilities has never been something they have spoken against.”

Regardless of whether Anthropic genuinely thinks it has a remotely realistic shot at pulling off a global pause — or if this is yet another ploy to boost its safety-minded image — it’s vowing to pursue further action.

“In the coming months, we will organize conversations where policymakers, researchers, civil society, and other AI companies can help answer some of the questions this piece raises, especially around full recursive self-improvement and how to create better options for coordination and deliberation,” the company wrote. “We’ll publish what comes out of it.” 

More on AI: CEO Says There Will Be No Raises Because He Spent All the Money on AI