惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Hugging Face - Blog
Hugging Face - Blog
Google DeepMind News
Google DeepMind News
云风的 BLOG
云风的 BLOG
WordPress大学
WordPress大学
Vercel News
Vercel News
Apple Machine Learning Research
Apple Machine Learning Research
T
Tailwind CSS Blog
I
InfoQ
小众软件
小众软件
Recent Announcements
Recent Announcements
博客园 - 【当耐特】
The GitHub Blog
The GitHub Blog
大猫的无限游戏
大猫的无限游戏
美团技术团队
T
The Blog of Author Tim Ferriss
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
酷 壳 – CoolShell
酷 壳 – CoolShell
MongoDB | Blog
MongoDB | Blog
V
V2EX
J
Java Code Geeks
有赞技术团队
有赞技术团队
博客园 - 聂微东
B
Blog RSS Feed
博客园 - 司徒正美

Cybersecurity Dive - Latest News

Dozens of Red Hat npm packages targeted in supply chain attack Turning tension into collaboration: How CIOs and CISOs can lead together Trump signs EO seeking early government access to powerful AI models Anthropic shares Mythos with 150 more organizations, including critical infrastructure operators Without strong governance, companies put credit ratings at risk in AI era CISA adds critical Palo Alto Networks firewall flaw to KEV as company, researchers warn of exploitation How Canva scaled to 260+M users while elevating security and productivity Top 4 data security best practices for the AI-enabled enterprise CISA urges security teams to check for software development compromises How CISOs can manage sovereign-cloud security risks IBM’s new $5B initiative will help enterprises rapidly patch open-source vulnerabilities Enterprise data is creeping its way into shadow AI tools Coordinated operation takes down Glassworm botnet Leading AI models are more vulnerable to malicious prompts than vendors claim Iranian government, not hacktivist group, breached LA Metro system, security firm says FBI warns about PhaaS platform used to access Microsoft 365 environments Iran-linked hackers target key US, allied sectors with sophisticated spear-phishing messages New York regulator calls for additional cyber mitigation amid heightened threat environment CISA asks cybersecurity community to alert it to vulnerability exploitation Grafana Labs links GitHub environment breach to TanStack npm supply chain attack 7-Eleven hit by data breach Microsoft disrupts cybercrime operation that hid behind legitimate software Compromised coding tool helped hackers breach thousands of GitHub repositories Telecom sector launches its own private ISAC Patch bypass allows hackers to exploit prior flaw in SonicWall SSL-VPN Grafana Labs says hacker gained access to codebase through leaked token How a government contest launched a revolution in AI-based bug hunting Attackers exploit critical flaw in Cisco Catalyst SD-WAN Controller MSPs need AI to fight AI-fueled cyberthreats: Guardz More money is going to physical security, but it’s often CISOs that oversee it: EY
NIST will test three major tech firms’ frontier AI models...
Eric Geller · 2026-05-06 · via Cybersecurity Dive - Latest News

An article from site logo

After Anthropic’s announcement of Claude Mythos, agencies across the government are racing to get ahead of new AI models’ potential dangers.

Published May 6, 2026

A large entrance sign that reads "Gate A, NIST, National Institute of Standards and Technology, U.S. Department of Commerce" is mounted on a rock base and surrounded by grass and trees. In the background to the left of the sign, there is a commercial building.

The front entrance sign at the Gaithersburg, Md., National Institute of Standards and Technology campus. R. Eskalis/NIST. Retrieved from NIST.

The U.S. government’s AI security center will evaluate frontier models from Google, Microsoft and xAI before their release to determine whether the models’ advanced capabilities pose cybersecurity risks.

The newly announced plan for the National Institute of Standards and Technology’s (NIST) Center for AI Standards and Innovation (CAISI) to conduct “pre-deployment evaluations” represents the U.S. government’s most significant attempt yet to get ahead of security threats from powerful AI systems.

“Independent, rigorous measurement science is essential to understanding frontier AI and its national security implications,” CAISI Director Chris Fall said in a statement. “These expanded industry collaborations help us scale our work in the public interest at a critical moment.”

NIST said the partnerships would help the agency and the tech companies exchange information, spur “voluntary product improvements” and ensure the government had a “clear understanding” of what AI models were capable of doing. An interagency task force at CAISI will allow officials from across the government to test the models, including in classified settings.

Natasha Crampton, Microsoft’s chief responsible AI officer, said in a LinkedIn post that tech companies couldn’t conduct “evaluations tied to national security and public safety” on their own.

“They require close collaboration between industry and governments with deep technical and security expertise,” she wrote, adding that Microsoft will apply what it learns “directly into how we design, test, and deploy AI — and share best practices to help strengthen AI testing more broadly.”

The arrangement represents a significant reversal for the Trump administration, which previously eliminated AI security review measures that it called overly burdensome.

The White House began rethinking its hands-off approach to AI after Anthropic announced that its latest model, Claude Mythos, was too dangerous to publicly release, because of its alarming ability to find serious software vulnerabilities. In addition to the new voluntary CAISI evaluations, the Trump administration is also considering instituting mandatory government reviews of all new AI models.

It remains unclear what testing standards CAISI will use for its evaluations. The outcomes that the agency defines as trustworthy and secure could be difficult to establish, according to Devin Lynch, a former director for cyber policy and strategy implementation at the White House Office of the National Cyber Director.

“Capability assessments are only as good as the threat models behind them,” Lynch wrote on LinkedIn. “CAISI will need to define, and publish, what it’s testing for, not just who it’s testing with.”