惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

V
V2EX
博客园 - 叶小钗
Last Week in AI
Last Week in AI
Google DeepMind News
Google DeepMind News
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
Microsoft Security Blog
Microsoft Security Blog
腾讯CDC
P
Proofpoint News Feed
大猫的无限游戏
大猫的无限游戏
The Cloudflare Blog
aimingoo的专栏
aimingoo的专栏
月光博客
月光博客
量子位
A
About on SuperTechFans
Engineering at Meta
Engineering at Meta
Apple Machine Learning Research
Apple Machine Learning Research
Jina AI
Jina AI
博客园 - Franky
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
人人都是产品经理
人人都是产品经理
D
DataBreaches.Net
博客园_首页
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
Stack Overflow Blog
Stack Overflow Blog

Opinion, Editorial, Views, Columnists, Columns | The HinduBusinessLine

Rupee can’t be defended from just one side Railways’ performance Why not have a women-only party? Labour pangs Pak’s peculiar comeback on the global stage Letters to Editor India has jobs, but it needs better ones Cross-border insolvency laws and trade A major health challenge Editorial. Snooping around Letters to the Editor dated April 20, 2026 Real-time metric for factory output All you want to know about the women’s reservation and delimitation bills fiasco Editorial. Process deficit Letters to the Editor dated April 19, 2026 WPI effect on new GDP series The tragic reality of police brutality India’s AI value paradox Prepare the ground India-Korea economic ties poised to strengthen Nari Shakti Bill — a missed opportunity Natural farming should become mainstream policy Insights from new GDP data Strategies to enhance fertilizer security Pathway to maritime insurance sovereignty Why the GoP’s jittery Clear the smoke Aiding piped gas push Stocks are the least over-priced asset in India Is TCS harassment case tip of the iceberg?
An AI model that’s too risky
2026-04-08 · via Opinion, Editorial, Views, Columnists, Columns | The HinduBusinessLine
Anthropic has decided who gets access to Claude Mythos Preview

Anthropic has decided who gets access to Claude Mythos Preview | Photo Credit: Dado Ruvic

Something unusual happened in the artificial intelligence industry this week. Anthropic, one of the leading AI labs, built a model so capable that it chose not to release it. The model, called Claude Mythos Preview, is not just another incremental advance. It can autonomously discover and exploit serious cybersecurity vulnerabilities — tasks that have historically required elite human researchers working for weeks. In one instance, it reportedly identified and exploited a long-standing remote code execution flaw in FreeBSD that allows an attacker to gain complete control over a server from anywhere on the internet, starting from an unauthenticated position. No human was involved in either the discovery or exploitation after the initial prompt.

That is not just a technical milestone. It is a glimpse of a near future in which AI systems could dramatically accelerate both cyber defence and cyber offence.

Pausing the launch

To its credit, Anthropic did something rare in Silicon Valley: it paused. Instead of launching the model, it created Project Glasswing, a coalition that includes Amazon Web Services, Apple, Broadcom, Cisco, CrowdStrike, Google, JPMorganChase, Microsoft and NVIDIA. The goal is to use the model defensively — to identify and patch vulnerabilities in critical systems before similar capabilities become widely available. It is hard to overstate how unusual this is. A company sat on what could be an enormously valuable commercial product because it judged the risks to global infrastructure too high. In an industry defined by rapid releases and competitive pressure, that decision deserves recognition.

But it should also make us uneasy. Because for all its promise, Project Glasswing exposes a deeper problem: the future of global cybersecurity may be shaped not by public institutions, but by the internal decisions of a handful of private companies.

Anthropic decided Claude Mythos Preview was too dangerous to release. It chose who would get access to it. It defined what counts as “defensive use.” And it will ultimately decide when systems with similar capabilities are safe enough for broader deployment.

That may be the right call. But it is still a private call. We have seen this pattern before in other high-stakes industries and rejected it. Banks do not determine their own capital requirements without oversight. Drug companies cannot unilaterally declare their products safe. Nuclear operators are not left to design their own inspection regimes. In each case, society concluded that even well-intentioned companies should not be the sole gatekeepers of technologies with systemic risk. Artificial intelligence is now at that point.

What would a more credible system look like? Start with independent verification. Today, when an AI company says a model is too dangerous or safe enough, there is no widely trusted external body that can rigorously audit that claim. That is a glaring gap.

Then consider coordination across the industry. Anthropic’s restraint matters little if a competitor races ahead. Systems from OpenAI, Google DeepMind or others may soon reach similar or greater capability levels. Without shared guardrails, the incentives to move fast will remain powerful.

There is also the question of representation. Project Glasswing is composed almost entirely of large corporations. They bring expertise and resources, but they do not represent everyone affected by these technologies. Small businesses, governments in the developing world and civil society groups have little say in how risks are prioritised or mitigated.

Finally, there is the issue of infrastructure. Some forms of safety, especially in cybersecurity, should not depend on the budgets or strategies of individual companies. They require sustained public investment and international cooperation.

The writer is a Distinguished Fellow at the Avinyum Foundation and former Managing Director of CGI India

Published on April 9, 2026