惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

博客园 - Franky
酷 壳 – CoolShell
酷 壳 – CoolShell
Google Online Security Blog
Google Online Security Blog
Engineering at Meta
Engineering at Meta
U
Unit 42
Security Latest
Security Latest
G
Google Developers Blog
www.infosecurity-magazine.com
www.infosecurity-magazine.com
D
Docker
T
Tailwind CSS Blog
Hacker News - Newest:
Hacker News - Newest: "LLM"
云风的 BLOG
云风的 BLOG
Hugging Face - Blog
Hugging Face - Blog
M
MIT News - Artificial intelligence
Threat Intelligence Blog | Flashpoint
Threat Intelligence Blog | Flashpoint
MongoDB | Blog
MongoDB | Blog
H
Help Net Security
Stack Overflow Blog
Stack Overflow Blog
C
Check Point Blog
S
Security Affairs
T
The Exploit Database - CXSecurity.com
S
SegmentFault 最新的问题
N
News and Events Feed by Topic
The GitHub Blog
The GitHub Blog
Apple Machine Learning Research
Apple Machine Learning Research
S
Securelist
IT之家
IT之家
P
Palo Alto Networks Blog
D
DataBreaches.Net
Help Net Security
Help Net Security
N
Netflix TechBlog - Medium
B
Blog RSS Feed
AWS News Blog
AWS News Blog
Scott Helme
Scott Helme
爱范儿
爱范儿
cs.AI updates on arXiv.org
cs.AI updates on arXiv.org
腾讯CDC
I
Intezer
J
Java Code Geeks
大猫的无限游戏
大猫的无限游戏
Microsoft Security Blog
Microsoft Security Blog
人人都是产品经理
人人都是产品经理
G
GRAHAM CLULEY
N
News | PayPal Newsroom
博客园 - 三生石上(FineUI控件)
A
Arctic Wolf
F
Fortinet All Blogs
The Register - Security
The Register - Security
Recent Commits to openclaw:main
Recent Commits to openclaw:main
博客园 - 【当耐特】

Opinion

Op-Ed: Microsoft’s July Patch Tuesday reveals 622 vulnerabilities Op-Ed: The transaction was legitimate; the crime was hidden in the system Op-Ed: Why CISOs are drowning in alerts but missing the real threat Op-Ed: The reality of data-centric security and attribute-based access control Op-Ed: Australia’s cyber law is stuck in the past – the Slay Review is our chance to fix it Australian federal budget 2026: The industry perspective Op-Ed: AI won’t patch the holes in your SOC Op-Ed: Australia inspired the EU’s online age restrictions, now it’s time for us to learn from them Op-Ed: Microsoft April Patch Tuesday reveals 167 vulnerabilities The industry speaks: World Identity Management Day 2026 Op-Ed: Why zero trust for OT should start at the boundary, not the boiler room The industry speaks: World Cloud Security Day 2026 The industry speaks: World Backup Day 2026 Op-Ed: Information sharing of cyber threats vital to national security Op-Ed: Building secure foundations for AI in the cloud Op-Ed: AI isn’t the threat, poor design is Op-Ed: Why Australia’s schools are becoming strategic targets for organised crime Op-Ed: Australia’s National AI Plan looks good on paper, but where are the teeth? Op-Ed: Australian organisations need federated authority to stay secure at scale
Op-Ed: Redefining performance in the AI-powered SOC
Simon Howe, Area Vice President, ANZ, ExtraHop · 2026-04-30 · via Opinion

Artificial intelligence is becoming the backbone of security operations, but measuring AI performance is a harder task than most CISOs realise.

AI is being embedded into every facet of the modern security operations centre (SOC), with the intention of reducing defensive loops and speeding up clearance of alert backlogs.

However, speed does not necessarily equal accuracy against highly sophisticated threats that understand how to evade detection.

You’re out of free articles for this month

To continue reading the rest of this article, please log in.

Without a backbone of relevant data, AI investments in the SOC will end up automating mistakes too fast for human analysts to intervene, rather than fulfilling the promise of machine-speed security. As a result, security and IT leaders must understand how to actually measure AI performance to build a resilient security posture.

The flaw in contemporary benchmarks

The current landscape of AI evaluation is held back by a significant gap: most frameworks measure what a system can do in a vacuum rather than how it actually performs under the pressure of a live breach. To bridge this divide, the industry relies on a few core metrics that translate raw processing power into operational value.

At the forefront is the mean time to detect (MTTD). Because AI can parse patterns across massive datasets far more efficiently than a human analyst, it fundamentally shrinks the window between an initial compromise and its discovery.

Once a threat is identified, the focus shifts to the mean time to respond (MTTR), where the speed at which AI automates the initial triage and suggests specific remediation steps is measured. AI workflows are also measured in the reduction of alerts, filtering out false positives, and prioritising the most critical threats.

While useful, these metrics possess two major flaws. They track volume and velocity but ignore whether the AI made the correct choice. Additionally, these figures are typically derived from a controlled laboratory setting rather than the high-pressure reality of a real-world attack.

AI security metrics that security leaders are missing

While MTTD and MTTR show that AI is working faster, they fail to reveal if the AI is fulfilling the promise of working smarter against sophisticated adversaries.

The adaptability gap: AI is typically trained on known threats. When attackers switch to living-off-the-land (LotL) techniques or adversarial prompts, there is a “relearning” phase. Traditional metrics don’t track how long it takes for an AI to recognise a pivot in attacker tactics, leaving a dangerous window where malicious activity is undetected and labelled as safe.

Shadow AI and unchecked logic: The unknown nature of AI often leads to a pipeline of unverified automated actions. When teams use unvetted large language models (LLMs) to write scripts or analyse logs, they introduce logic that has never been audited, creating hidden vulnerabilities.

The erosion of human oversight: In the pursuit of high-speed processes, the human element, still very much critical in this current stage of AI development, is often treated as a bottleneck. If analysts stop questioning AI verdicts to keep their metrics high, the system becomes a single point of failure.

Strategies for authentic AI evaluation

To truly measure success, security leaders must move away from controlled testing and towards testing founded in reality:

Stress test under adversity: Evaluate AI performance using degraded data, simultaneous high-priority alerts, and extreme time constraints. True effectiveness is measured by how well the AI prioritises threats when the system is saturated, and telemetry is incomplete.

Map the failure points: Organisations must identify exactly where an AI’s confidence begins to drop. By mapping these boundaries, teams can build a reliable hand-off process, ensuring that low-confidence AI decisions are automatically routed to human experts for validation.

Demand transparent evidence: Instead of trusting a binary verdict of pass or fail, use tools that expose the AI’s reasoning. Teams need to see which signals were prioritised, which behavioural indicators were flagged, and most importantly, which data points the model chose to ignore.

Adopting AI based on operational reality rather than marketing hype ensures a more resilient defence.

By closing the measurement gap and using solutions based on visibility to actually observe what AI is doing within the SOC, organisations can move beyond assumptions and build a security posture that is truly accountable, defensible, and capable of withstanding the scrutiny of a real-world attack.

Cyber DailyWant to see more stories from trusted news sources?
Make Cyber Daily a preferred news source on Google.