惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

V
Visual Studio Blog
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
N
Netflix TechBlog - Medium
博客园 - 叶小钗
大猫的无限游戏
大猫的无限游戏
S
SegmentFault 最新的问题
V
V2EX
IT之家
IT之家
J
Java Code Geeks
Hacker News - Newest:
Hacker News - Newest: "LLM"
cs.AI updates on arXiv.org
cs.AI updates on arXiv.org
GbyAI
GbyAI
D
Docker
S
Secure Thoughts
Recent Announcements
Recent Announcements
Webroot Blog
Webroot Blog
Application and Cybersecurity Blog
Application and Cybersecurity Blog
云风的 BLOG
云风的 BLOG
博客园_首页
cs.CV updates on arXiv.org
cs.CV updates on arXiv.org
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
Security Archives - TechRepublic
Security Archives - TechRepublic
酷 壳 – CoolShell
酷 壳 – CoolShell
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
N
News | PayPal Newsroom
S
Security @ Cisco Blogs
I
InfoQ
Last Week in AI
Last Week in AI
SecWiki News
SecWiki News
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
W
WeLiveSecurity
T
Troy Hunt's Blog
Recent Commits to openclaw:main
Recent Commits to openclaw:main
H
Hackread – Cybersecurity News, Data Breaches, AI and More
Attack and Defense Labs
Attack and Defense Labs
美团技术团队
T
The Blog of Author Tim Ferriss
Google DeepMind News
Google DeepMind News
Martin Fowler
Martin Fowler
B
Blog
CTFtime.org: upcoming CTF events
CTFtime.org: upcoming CTF events
Scott Helme
Scott Helme
T
Tor Project blog
Know Your Adversary
Know Your Adversary
有赞技术团队
有赞技术团队
Hugging Face - Blog
Hugging Face - Blog
Recorded Future
Recorded Future
C
Cyber Attacks, Cyber Crime and Cyber Security
AI
AI
G
Google Developers Blog

Opinion

Op-Ed: Microsoft’s July Patch Tuesday reveals 622 vulnerabilities Op-Ed: The transaction was legitimate; the crime was hidden in the system Op-Ed: Why CISOs are drowning in alerts but missing the real threat Op-Ed: The reality of data-centric security and attribute-based access control Op-Ed: Australia’s cyber law is stuck in the past – the Slay Review is our chance to fix it Australian federal budget 2026: The industry perspective Op-Ed: AI won’t patch the holes in your SOC Op-Ed: Australia inspired the EU’s online age restrictions, now it’s time for us to learn from them Op-Ed: Microsoft April Patch Tuesday reveals 167 vulnerabilities The industry speaks: World Identity Management Day 2026 Op-Ed: Why zero trust for OT should start at the boundary, not the boiler room The industry speaks: World Cloud Security Day 2026 The industry speaks: World Backup Day 2026 Op-Ed: Information sharing of cyber threats vital to national security Op-Ed: Building secure foundations for AI in the cloud Op-Ed: AI isn’t the threat, poor design is Op-Ed: Why Australia’s schools are becoming strategic targets for organised crime Op-Ed: Australia’s National AI Plan looks good on paper, but where are the teeth? Op-Ed: Australian organisations need federated authority to stay secure at scale Supply chain risk: Understanding the weakest link in cyber security
Op-Ed: Redefining performance in the AI-powered SOC
Simon Howe, Area Vice President, ANZ, ExtraHop · 2026-04-30 · via Opinion

Artificial intelligence is becoming the backbone of security operations, but measuring AI performance is a harder task than most CISOs realise.

AI is being embedded into every facet of the modern security operations centre (SOC), with the intention of reducing defensive loops and speeding up clearance of alert backlogs.

However, speed does not necessarily equal accuracy against highly sophisticated threats that understand how to evade detection.

You’re out of free articles for this month

To continue reading the rest of this article, please log in.

Without a backbone of relevant data, AI investments in the SOC will end up automating mistakes too fast for human analysts to intervene, rather than fulfilling the promise of machine-speed security. As a result, security and IT leaders must understand how to actually measure AI performance to build a resilient security posture.

The flaw in contemporary benchmarks

The current landscape of AI evaluation is held back by a significant gap: most frameworks measure what a system can do in a vacuum rather than how it actually performs under the pressure of a live breach. To bridge this divide, the industry relies on a few core metrics that translate raw processing power into operational value.

At the forefront is the mean time to detect (MTTD). Because AI can parse patterns across massive datasets far more efficiently than a human analyst, it fundamentally shrinks the window between an initial compromise and its discovery.

Once a threat is identified, the focus shifts to the mean time to respond (MTTR), where the speed at which AI automates the initial triage and suggests specific remediation steps is measured. AI workflows are also measured in the reduction of alerts, filtering out false positives, and prioritising the most critical threats.

While useful, these metrics possess two major flaws. They track volume and velocity but ignore whether the AI made the correct choice. Additionally, these figures are typically derived from a controlled laboratory setting rather than the high-pressure reality of a real-world attack.

AI security metrics that security leaders are missing

While MTTD and MTTR show that AI is working faster, they fail to reveal if the AI is fulfilling the promise of working smarter against sophisticated adversaries.

The adaptability gap: AI is typically trained on known threats. When attackers switch to living-off-the-land (LotL) techniques or adversarial prompts, there is a “relearning” phase. Traditional metrics don’t track how long it takes for an AI to recognise a pivot in attacker tactics, leaving a dangerous window where malicious activity is undetected and labelled as safe.

Shadow AI and unchecked logic: The unknown nature of AI often leads to a pipeline of unverified automated actions. When teams use unvetted large language models (LLMs) to write scripts or analyse logs, they introduce logic that has never been audited, creating hidden vulnerabilities.

The erosion of human oversight: In the pursuit of high-speed processes, the human element, still very much critical in this current stage of AI development, is often treated as a bottleneck. If analysts stop questioning AI verdicts to keep their metrics high, the system becomes a single point of failure.

Strategies for authentic AI evaluation

To truly measure success, security leaders must move away from controlled testing and towards testing founded in reality:

Stress test under adversity: Evaluate AI performance using degraded data, simultaneous high-priority alerts, and extreme time constraints. True effectiveness is measured by how well the AI prioritises threats when the system is saturated, and telemetry is incomplete.

Map the failure points: Organisations must identify exactly where an AI’s confidence begins to drop. By mapping these boundaries, teams can build a reliable hand-off process, ensuring that low-confidence AI decisions are automatically routed to human experts for validation.

Demand transparent evidence: Instead of trusting a binary verdict of pass or fail, use tools that expose the AI’s reasoning. Teams need to see which signals were prioritised, which behavioural indicators were flagged, and most importantly, which data points the model chose to ignore.

Adopting AI based on operational reality rather than marketing hype ensures a more resilient defence.

By closing the measurement gap and using solutions based on visibility to actually observe what AI is doing within the SOC, organisations can move beyond assumptions and build a security posture that is truly accountable, defensible, and capable of withstanding the scrutiny of a real-world attack.

Cyber DailyWant to see more stories from trusted news sources?
Make Cyber Daily a preferred news source on Google.