惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

量子位
博客园_首页
罗磊的独立博客
云风的 BLOG
云风的 BLOG
J
Java Code Geeks
Last Week in AI
Last Week in AI
D
DataBreaches.Net
Jina AI
Jina AI
博客园 - Franky
大猫的无限游戏
大猫的无限游戏
Apple Machine Learning Research
Apple Machine Learning Research
V
V2EX
D
Docker
MongoDB | Blog
MongoDB | Blog
B
Blog RSS Feed
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
宝玉的分享
宝玉的分享
Engineering at Meta
Engineering at Meta
The Cloudflare Blog
博客园 - 三生石上(FineUI控件)
有赞技术团队
有赞技术团队
人人都是产品经理
人人都是产品经理
H
Help Net Security
T
The Blog of Author Tim Ferriss

SiliconANGLE

Will agentic AI governance run amok? The lesson of Asimov’s Three Laws - SiliconANGLE AI + quantum, Amazon vs. Starlink and the wide-open US-China internet battle - SiliconANGLE Team Cymru launches Total Insights Feed to replace legacy threat intelligence lists - SiliconANGLE AI Mode in Chrome adds split-screen view to enhance the web search experience - SiliconANGLE Resolve AI raises $40M at $1.5B valuation to optimize production environments - SiliconANGLE How Zscaler and OpenAI turn zero-trust security into an AI accelerator - SiliconANGLE OpenAI ratchets up Codex's agentic capabilities to rival Claude Code - SiliconANGLE Anthropic launches Claude Opus 4.7 with coding, visual reasoning improvements - SiliconANGLE Slash raises $100M at a $1.4B valuation to expand AI-powered banking platform for online businesses - SiliconANGLE Canva unveils Canva AI 2.0, recasting its platform as an agentic system for work - SiliconANGLE Data center, consumer device chips boost TSMC’s revenue - SiliconANGLE Mission-critical security cannot be bolted on, says Oracle - SiliconANGLE Agentic infrastructure reshapes enterprise AI - SiliconANGLE Data quality, and data freedom, foundational for AI success - SiliconANGLE Data trust is a bedrock in successful, scalable AI outcomes - SiliconANGLE Google introduces new agentic AI-ready tools and resources for Android developers  - SiliconANGLE Agentic AI orchestration separates winners from laggards - SiliconANGLE Data-driven tools turning the tide against human trafficking - SiliconANGLE Achieving trusted AI development goes beyond 'vibes' - SiliconANGLE Impinj boosts edge computing power in updated R700 RAIN RFID reader - SiliconANGLE Certinia powers professional services with AI - SiliconANGLE Antioch prepares to accelerate simulated testing for autonomous robots after raising $8.5M - SiliconANGLE Developer tooling startup Expo nabs $45M investment - SiliconANGLE Solidroad lands $25M to bring AI to customer support interactions - SiliconANGLE DuploCloud lands compliance and AI governance certifications as enterprise buyers tighten scrutiny - SiliconANGLE Lua lands $5.8M to help businesses build and manage AI agent workforces - SiliconANGLE Best of frenemies: Oracle's and AWS' clouds unite with dedicated, private connectivity - SiliconANGLE NIST shifts National Vulnerability Database to risk-based triage as CVE submissions hit record levels - SiliconANGLE Cisco goes to the races with new Churchill Downs multiyear partnership - SiliconANGLE Susecon 2026 will tackle the future of open-source platforms - SiliconANGLE
AI constraints must come before deployment, not after - S...
John Waller · 2026-05-02 · via SiliconANGLE

AI constraints must come before deployment, not after

On April 7, 2026, Anthropic did something unprecedented in the history of artificial intelligence: The company announced that it had built its most capable model ever and would not be releasing it to the public.

The model had not failed. In fact, it had performed so well, across such consequential domains, that Anthropic concluded the constraint infrastructure required to deploy it responsibly did not yet exist.

In the weeks of testing before the announcement, Claude Mythos Preview had identified critical vulnerabilities in every major operating system and every major web browser – thousands of flaws that had survived, in some cases, decades of human review and millions of automated security tests. The same capability that made it an extraordinary defensive tool made it, in the wrong hands, a means to compromise virtually any major software system in the world.

Anthropic’s response was Project Glasswing: a consortium of 50 of the leading technology and critical infrastructure organizations committed to finding and patching vulnerabilities before the capability proliferated beyond responsible actors. The company was explicit about why Mythos itself would remain unreleased: “We need to make progress in developing cybersecurity and other safeguards that detect and block the model’s most dangerous outputs.” The most safety-focused AI laboratory in the world had built a system it could not yet safely constrain, so it paused.

For many organizations deploying AI, that question comes later – if it comes at all.

What AI missing by design

Human beings do not require external governance to prevent the most harmful behaviors. We are constrained from within by biology, social accountability, legal consequence and the cognitive limits that prevent any individual from optimizing at machine speed and scale. These constraints were not designed; they emerged over millennia. They are imperfect, but they exist as a baseline.

AI systems inherit none of these. Every limit is one someone chose to engineer. An AI system given an objective will pursue it through whatever path is mathematically available – including those that involve collusion, discriminatory outcomes, unauthorized resource acquisition, or, as Mythos Preview demonstrated, the autonomous exploitation of critical infrastructure vulnerabilities. It’s not because the system is malicious, but because nothing was in place to prevent it.

This is not a flaw. It is the nature of these systems, and it is the central governance challenge every organization deploying AI faces today.

Governance is a work in progress

A mature AI governance program looks like other rigorous organizational disciplines such as DevSecOps, regulatory compliance and financial controls. It inventories every AI system in production, assesses it against a proportional set of technical, operational and governance controls, measures the gap between what is prescribed and what is actually implemented and reviews that gap on a defined schedule as systems and their environments evolve. It is systematic, documented and auditable – not a policy document, but a practice.

That standard exists in other domains because those domains built it over decades of incidents, regulation, and accumulated institutional knowledge. AI governance is only a few years into that same process. Most organizations have not yet had the time, the mandate or the forcing function to develop their AI governance to the same level of rigor as the compliance and security practices they have spent years maturing.

Competitive pressure compounds the problem. With so much market uncertainty and a regulatory environment still taking shape, many organizations are moving faster than their governance programs can keep pace. We are seeing the maturity of the industry, its standards and its regulations built in real time.

Constraints must come first

The most important lesson of the Glasswing announcement is about sequence. Anthropic did not build Mythos Preview and then ask whether it was safe to release. The company evaluated the system’s capabilities rigorously, concluded that the constraint infrastructure didn’t exist to deploy it responsibly and chose to withhold it from the public. The governance question came before the deployment decision.

Unfortunately, that sequence is more often the exception than the rule in businesses, due to market forces that reward speed and a governance ecosystem that has not yet caught up.

Writing on the day of the announcement, New York Times columnist Thomas Friedman called what Mythos Preview represents potentially as consequential as the emergence of nuclear weapons and the need for nonproliferation, a capability no single organization or country can manage alone. He is not wrong, but the civilizational scale does not excuse the organizational one. Every organization deploying AI systems today faces a version of the same question Anthropic answered with Mythos: Is the constraint infrastructure adequate relative to the capability being deployed?

Many organizations do not yet have a reliable answer. That’s not from indifference, but because the frameworks, standards and regulatory guidance needed to make that evaluation with confidence are still being developed.

Don’t defer the decision

Project Glasswing is a beginning, involving multiple organizations, a defensive mandate and a $100 million commitment applied to a specific threat. It is not a solution to the broader challenge it has illuminated.

That challenge belongs to every organization that builds or deploys AI. Treat constraint adequacy as a deployment prerequisite, not a post-deployment remediation task. Measure the gap between what governance documents say and what AI systems actually do. Recognize that as AI capability advances, the constraint systems designed for current capabilities require continuous reassessment.

Anthropic’s choice demonstrated something rare: the discipline to ask the governance question honestly and act on the answer, even when the answer was inconvenient.

The organizations that will be on the right side of AI’s history are the ones asking that question now – before the incident that makes the answer undeniable.

John Waller is risk advisory practice lead at managed security services firm UltraViolet Cyber Inc. He wrote this article for SiliconANGLE.

Image: Wikimedia Commons

A message from John Furrier, co-founder of SiliconANGLE:

Support our mission to keep content open and free by engaging with theCUBE community. Join theCUBE’s Alumni Trust Network, where technology leaders connect, share intelligence and create opportunities.

  • 15M+ viewers of theCUBE videos, powering conversations across AI, cloud, cybersecurity and more
  • 11.4k+ theCUBE alumni — Connect with more than 11,400 tech and business leaders shaping the future through a unique trusted-based network.

About SiliconANGLE Media

SiliconANGLE Media is a recognized leader in digital media innovation, uniting breakthrough technology, strategic insights and real-time audience engagement. As the parent company of SiliconANGLE, theCUBE Network, theCUBE Research, CUBE365, theCUBE AI and theCUBE SuperStudios — with flagship locations in Silicon Valley and the New York Stock Exchange — SiliconANGLE Media operates at the intersection of media, technology and AI.

Founded by tech visionaries John Furrier and Dave Vellante, SiliconANGLE Media has built a dynamic ecosystem of industry-leading digital media brands that reach 15+ million elite tech professionals. Our new proprietary theCUBE AI Video Cloud is breaking ground in audience interaction, leveraging theCUBEai.com neural network to help technology companies make data-driven decisions and stay at the forefront of industry conversations.