惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Recent Announcements
Recent Announcements
Martin Fowler
Martin Fowler
MongoDB | Blog
MongoDB | Blog
Engineering at Meta
Engineering at Meta
Stack Overflow Blog
Stack Overflow Blog
Google DeepMind News
Google DeepMind News
Microsoft Security Blog
Microsoft Security Blog
aimingoo的专栏
aimingoo的专栏
I
InfoQ
B
Blog
WordPress大学
WordPress大学
Jina AI
Jina AI
小众软件
小众软件
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
博客园_首页
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
酷 壳 – CoolShell
酷 壳 – CoolShell
阮一峰的网络日志
阮一峰的网络日志
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
G
Google Developers Blog
C
Check Point Blog
月光博客
月光博客
L
LangChain Blog
GbyAI
GbyAI

Security

Report: Business email compromise attacks surged dangerously in April Scope Systems confirms cyber incident, says no data loss occurred Instructure breach: ShinyHunters says ‘matter has been resolved’ Rapid7 launches Cyber GRC program to connect compliance with live risk data Australian federal budget 2026: The industry perspective Op-Ed: Microsoft May Patch Tuesday reveals 137 vulnerabilities Federal Budget 2026: The state of cyber security spending for the coming year OpenAI offers EU early access to its cyber security model Exclusive: Aussie firm Earth Systems listed by INC Ransom hacking group Op-Ed: Why Middle East tensions demand immediate action on OT security Aussie schools breach: Instructure boss “reaches agreement” with ShinyHunters to not release data Institute of Public Accountants members hit by data breach Union demands answers on Qantas AI plans 1 in 3 small businesses don't think they're a cyber target, new research finds Exclusive: Aussie toy distributor listed by M3rx ransomware Exclusive: Australian Computer Society investigating possible breach after ShinyHunters hack claims The industry speaks – part 2: World Password Day 2026 Aussie schools breach: The Instructure hack “transcends an isolated IT incident” Exclusive: Aussie car part importer Strategic Imports allegedly breached by threat actors New South Wales, other states, investigating Instructure/Canvas data breach Australian Cyber Security Centre warns of ClickFix campaign leveraging Australian infrastructure Queensland Department of Education confirms students & staff impacted by ShinyHunters data breach ACMA takes action against SpinTel & Yomojo over mobile number fraud violations The Industry Speaks, Part 1: World Password Day 2026 Qualys and Converge tie cyber insurance pricing to real-time security posture Fakeout: Iranian APT caught hiding behind Chaos ransomware activity Exclusive: Australian energy management firm allegedly breached by SafePay Real estate giant Cushman & Wakefield confirms cyber incident, Qilin and ShinyHunters claim attack CrowdStrike expands Project QuiltWorks as more partners join AI security coalition Hacked: ALS discloses cyber incident, unauthorised access to IT systems
Op-Ed: Redefining performance in the AI-powered SOC
Simon Howe, Area Vice President, ANZ, ExtraHop · 2026-04-30 · via Security

Artificial intelligence is becoming the backbone of security operations, but measuring AI performance is a harder task than most CISOs realise.

AI is being embedded into every facet of the modern security operations centre (SOC), with the intention of reducing defensive loops and speeding up clearance of alert backlogs.

However, speed does not necessarily equal accuracy against highly sophisticated threats that understand how to evade detection.

You’re out of free articles for this month

To continue reading the rest of this article, please log in.

Without a backbone of relevant data, AI investments in the SOC will end up automating mistakes too fast for human analysts to intervene, rather than fulfilling the promise of machine-speed security. As a result, security and IT leaders must understand how to actually measure AI performance to build a resilient security posture.

The flaw in contemporary benchmarks

The current landscape of AI evaluation is held back by a significant gap: most frameworks measure what a system can do in a vacuum rather than how it actually performs under the pressure of a live breach. To bridge this divide, the industry relies on a few core metrics that translate raw processing power into operational value.

At the forefront is the mean time to detect (MTTD). Because AI can parse patterns across massive datasets far more efficiently than a human analyst, it fundamentally shrinks the window between an initial compromise and its discovery.

Once a threat is identified, the focus shifts to the mean time to respond (MTTR), where the speed at which AI automates the initial triage and suggests specific remediation steps is measured. AI workflows are also measured in the reduction of alerts, filtering out false positives, and prioritising the most critical threats.

While useful, these metrics possess two major flaws. They track volume and velocity but ignore whether the AI made the correct choice. Additionally, these figures are typically derived from a controlled laboratory setting rather than the high-pressure reality of a real-world attack.

AI security metrics that security leaders are missing

While MTTD and MTTR show that AI is working faster, they fail to reveal if the AI is fulfilling the promise of working smarter against sophisticated adversaries.

The adaptability gap: AI is typically trained on known threats. When attackers switch to living-off-the-land (LotL) techniques or adversarial prompts, there is a “relearning” phase. Traditional metrics don’t track how long it takes for an AI to recognise a pivot in attacker tactics, leaving a dangerous window where malicious activity is undetected and labelled as safe.

Shadow AI and unchecked logic: The unknown nature of AI often leads to a pipeline of unverified automated actions. When teams use unvetted large language models (LLMs) to write scripts or analyse logs, they introduce logic that has never been audited, creating hidden vulnerabilities.

The erosion of human oversight: In the pursuit of high-speed processes, the human element, still very much critical in this current stage of AI development, is often treated as a bottleneck. If analysts stop questioning AI verdicts to keep their metrics high, the system becomes a single point of failure.

Strategies for authentic AI evaluation

To truly measure success, security leaders must move away from controlled testing and towards testing founded in reality:

Stress test under adversity: Evaluate AI performance using degraded data, simultaneous high-priority alerts, and extreme time constraints. True effectiveness is measured by how well the AI prioritises threats when the system is saturated, and telemetry is incomplete.

Map the failure points: Organisations must identify exactly where an AI’s confidence begins to drop. By mapping these boundaries, teams can build a reliable hand-off process, ensuring that low-confidence AI decisions are automatically routed to human experts for validation.

Demand transparent evidence: Instead of trusting a binary verdict of pass or fail, use tools that expose the AI’s reasoning. Teams need to see which signals were prioritised, which behavioural indicators were flagged, and most importantly, which data points the model chose to ignore.

Adopting AI based on operational reality rather than marketing hype ensures a more resilient defence.

By closing the measurement gap and using solutions based on visibility to actually observe what AI is doing within the SOC, organisations can move beyond assumptions and build a security posture that is truly accountable, defensible, and capable of withstanding the scrutiny of a real-world attack.

Cyber DailyWant to see more stories from trusted news sources?
Make Cyber Daily a preferred news source on Google.