惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Microsoft Security Blog
Microsoft Security Blog
J
Java Code Geeks
GbyAI
GbyAI
aimingoo的专栏
aimingoo的专栏
L
LangChain Blog
I
InfoQ
D
Docker
F
Fortinet All Blogs
Y
Y Combinator Blog
Martin Fowler
Martin Fowler
月光博客
月光博客
B
Blog
Engineering at Meta
Engineering at Meta
T
Tailwind CSS Blog
罗磊的独立博客
博客园_首页
G
Google Developers Blog
Stack Overflow Blog
Stack Overflow Blog
Recent Announcements
Recent Announcements
D
DataBreaches.Net
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
B
Blog RSS Feed
IT之家
IT之家
V
V2EX

Cisco Blogs

Edge opportunity for service providers: Turn infrastructure into new services MRC and SRv6: How Foundational Networking Innovations Are Enabling the Next Generation of AI Supercomputers The SMB Marketing Reset: Winning Customer Trust in a Digital-First Economy Inside the SOC: AI-powered DNS defense against ransomware Our Path Forward Securing the Federal Digital Experience with Cisco ThousandEyes for Government Cisco at ONUG Dallas 2026: Securing the AI Data Center in the Agentic Era Cisco and Red Hat are powering intelligent core to edge: Red Hat Summit insights Building the Capabilities That Win: How Cisco Partners Can Lead in the SMB & Mid-Market Era How Two Hours Felt Bigger Than My To-Do List Announcing Foundry Security Spec Ace the CCIE Collaboration Lab: Success Tips from a TAC Engineer Turned CCIE Protecting Agents with Cisco AI Defense and Google Agent Development Kit Powering an Inclusive Future: Your guide to the Purpose Pavilion at Cisco Live Las Vegas The Infrastructure Behind the Mission: SOF Week 2026 Cisco Networking App Marketplace Partners at Cisco Live 2026 Beyond the Pilot: Building the Clinical Data Fabric for the Agentic Era Benchmarking scale-out AI fabrics with Cisco N9000 + AMD Pensando™ Pollara 400 NICs Month of Developer Productivity: Build and Forget The race to autonomous transport networks: A new study Lean IT, future-ready: How to save time and simplify wireless management with AI Reading Between the Pixels: Failure Modes in Vision Language Models Biochar’s triple win: Healthier soils, improved crops, and decarbonization Designing a Proactive Customer Journey Modernize your data center operations with Cisco Nexus Dashboard Why your automation stack needs Cisco Agentic Workflows Try Cisco AI Defense Explorer Edition in this hands-on lab From Bandwidth to Intelligence: How Cisco is Powering AI-Ready Networks Spotlight on digital transformation | FY25 Purpose Report Galaxy Mode is live: A limited-time look at what your Cisco AI Assistant and AgenticOps can already do
Maximizing Uptime: The Power of AI Troubleshooting for In...
Rohit Agarwalla · 2026-06-25 · via Cisco Blogs

Industrial environments are entering the era of Physical AI. Driven by machine vision, autonomous vehicles, and Software-Defined Automation, this new intelligence sits on top of thousands of already-networked PLCs, HMIs, safety controllers, and motor drives. Because every piece of the factory floor is now hyper-connected, maximizing network uptime is no longer optional—it is a critical business mandate. 

While network anomalies are unavoidable, effective troubleshooting is essential to minimizing mean time to detection (MTTD) and resolution (MTTR).

The industrial network troubleshooting gap 

  • Current approaches are slow for the factory floor. When an issue disrupts production, every minute counts. But today’s troubleshooting is largely reactive – problems surface when a line stops or a device goes unreachable, and then the investigation begins. Correlating issues to root cause is manual, spread across multiple tools, and depends on whoever happens to be available. In an environment where downtime is measured in tens of thousands of dollars per minute, that process doesn’t move fast enough. 
  • Too many escalations for too few experts. The first responder – the maintenance technician on the floor — knows the physical systems but struggles to diagnose when an issue is network-related. IT tools lack enough OT context to help, and OT technicians lack networking expertise to use these tools. Even straightforward problems – for example, an OT endpoint that was accidentally moved to a different port causing it to go offline – get escalated because the first responder is unable to determine the root cause. The OT escalation point – the network expert team that absorb these escalations is small and stretched across sites. 

The result: hours of production downtime while experts catch up. For physical-layer issues – a damaged cable, a failing fiber optic transceiver – the fix is often simple enough for the technician on the floor to act on directly, if they can get to root cause. For network operations issues, it still needs the network experts – but the gap is the same: getting from issue to root cause fast enough to keep the line moving.

Figure 1: Most network issues need escalation to experts wasting precious time

A digital teammate for your OT team 

As part of Cisco AgenticOps and available through Cisco Cloud Control, AI Troubleshooting for Industrial Networks is an always-on ambient agent in the factory floor that acts as a digital teammate for your OT team – giving technicians a path from symptoms to root cause, and giving network engineers a headstart when they need to step in. 

The on-premises, ambient agent senses the environment 24×7, detects alerts and patterns, diagnoses the signals, and prepares recommended actions before a maintenance technician has to ask. It detects issues by monitoring switch system messages and clustering related events in a time window — rather than treating every alert as a separate incident. It diagnoses root causes using deterministic logic built on Cisco’s industrial networking expertise. By gathering and reasoning over evidence from the network’s topology, state and configuration, the agent quickly identifies the most likely cause. And then it recommends clear, sequenced next steps – whether that’s a physical fix the OT technician can follow or a precise escalation for a network configuration issue the network expert can act on immediately. 

An example: A machine in the packing area suddenly halts. The agent detects a problem with the fiber connection from the access switch, gathers interface and SFP state, and determines that the SFP on port 1/1 is experiencing signal degradation, likely due to environmental dust blocking the signal. The alert tells the OT technician exactly which switch and port are affected and provides a clear physical fix: clean and reseat the SFP module. Without the agent, this same issue would have been reported as “comms fault” by the OT technician, escalated to the network expert team, and diagnosed hours later. 

Figure 2: The intuitive agent interface displays detected issues, root causes, actionable fixes, and the affected network topology

The agent handles the most common issues experienced on the factory floor – spanning physical faults and operational disruptions – through the evidence-driven diagnostic logic: 

  • Cable and fiber optic faults: Detects link instability and determines whether the cause is physical such as a damaged cable or fiber optic module. For suspected cable damage, it can run a cable diagnostic test (with technician consent) to pinpoint the fault distance from the switch. 
  • Endpoint device offline: Investigates non-physical reasons why an endpoint stopped communicating such as duplex mismatch, endpoint moved to a different switch port with VLAN mismatch or duplicate IP due to L2NAT misconfiguration.  
  • Power over Ethernet (PoE) failures: Checks power delivery status, available budget, recent power events, and enforcement status to determine whether the cause is a port-level policy fault or insufficient switch power budget.
  • Switch power supply failures: Monitors for power supply failure, input power quality, surfaces the loss of a redundant power supply. 
  • Switch stability issues: Monitors high memory or CPU utilization, warns a process is consuming up CPU cycles, enabling technicians to escalate with diagnostic data.

Everyday operational questions

Beyond proactive alerting, the agent helps OT teams answer common questions without needing to log into a switch and run CLI commands. OT teams can select a switch and start a conversation with it to get live operational and configuration data. The agent also suggests the most relevant prompts based on the device and context.  Network experts can tag devices with familiar names, locations, and production areas (e.g., “Line 1 welder”), so OT teams can query switches using OT language instead of IP addresses or hostnames.

Figure 1: Equipped with the AI agent, first responders can resolve most network cases on their own, saving critical time and reducing escalations.

As one customer OT network expert from an early alpha trial put it: “This will help me sleep better at night — it’ll reduce escalations during testing and bring up.” AI Troubleshooting for Industrial Networks is designed to close the gap between symptoms and root causes on the factory floor — reducing escalations, compressing resolution times, and keeping production moving.  

The promise of Physical AI relies entirely on maximizing network uptime. AI Troubleshooting for Industrial Networks empowers your OT teams to slash downtime and secure the foundation for this new era.

If you are interested in shaping the next phase of the agent and gaining access, join the beta program today. 

Learn more

At-a-glance overview

Connect with our manufacturing experts

Authors

Avatar

Rohit Agarwalla

Senior Director Product Management and Technical Marketing

Enterprise Networking

Explore Cisco Industrial IoT

Improve business outcomes with our end-to-end Industrial IoT solutions. Securely connect assets, applications, and data in real time to apply transformative business changes.