惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Microsoft Azure Blog
Microsoft Azure Blog
Engineering at Meta
Engineering at Meta
A
About on SuperTechFans
T
The Blog of Author Tim Ferriss
I
InfoQ
博客园_首页
G
Google Developers Blog
爱范儿
爱范儿
Last Week in AI
Last Week in AI
量子位
阮一峰的网络日志
阮一峰的网络日志
雷峰网
雷峰网
酷 壳 – CoolShell
酷 壳 – CoolShell
Vercel News
Vercel News
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
GbyAI
GbyAI
月光博客
月光博客
The GitHub Blog
The GitHub Blog
V
Visual Studio Blog
N
Netflix TechBlog - Medium
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
博客园 - 司徒正美
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
博客园 - 聂微东

Datadog | The Monitor blog

Introducing our open source AI-native SAST Instrument and monitor Boomi integration flows with OpenTelemetry and Datadog Not all index scans are equal: How we cut query latency by over 99% Platform engineering metrics: What to measure and what to ignore Integrate Recorded Future threat intelligence with Datadog Cloud SIEM CI/CD security: threat modeling using a MITRE-style threat matrix CI/CD security: How to secure your GitHub ecosystem Ingress NGINX is EOL: A practical guide for migrating to Kubernetes Gateway API Operating agentic AI with Amazon Bedrock AgentCore and Datadog LLM Observability: Lessons from NTT DATA Introducing the Datadog Code Security MCP Capture and analyze custom heatmaps in Session Replay Understand session replays faster with AI summaries and smart chapters Monitor ClickHouse query performance with Datadog Database Monitoring How we designed empathetic alert sounds for on-call engineers Search and act across Datadog to resolve issues faster with Bits Assistant Measure the business impact of every product change with Datadog Experiments Analyzing round trip query latency Configuring JavaScript caches for better performance Introducing Bits AI Dev Agent for Code Security Datadog achieves ISO 42001 certification for responsible AI Monitor Nutanix clusters, hosts, and VMs with Datadog Monitor Juniper Mist in Datadog A new Host Map for modern infrastructure Annotate traces to improve LLM quality with Datadog LLM Observability What’s new in Cloud SIEM: AI-powered investigations, enhanced threat intelligence, and scalable security operations Explore Kubernetes with native OpenTelemetry data Monitor Oracle Fusion Cloud Applications with Datadog Announcing the Datadog Terraform provider v4.0.0 Scaling Kubernetes workloads on custom metrics How to design cloud environments for AI-powered threat analysis
Streamline network investigations with an enhanced queryi...
2023-05-08 · via Datadog | The Monitor blog

Editor’s note: This post covers Cloud Network Monitoring, a Datadog feature that was originally called Network Performance Monitoring.

Effective network troubleshooting requires collecting and correlating thousands of data points across your entire stack. The more data you ingest, however, the more data you have to search through in order to locate important signals. This can make it hard to find the information you need during time-sensitive investigations.

Datadog Network Performance Monitoring (NPM) provides enhanced search and mapping capabilities that help you quickly pinpoint critical metrics, giving you clear visibility into the health of your network at any given time. NPM’s new search experience delivers streamlined, unified filtering that enables you to comb through both client and server data simultaneously for a complete picture of communication across your network. And with clustered maps, you can organize large quantities of network data into neatly categorized visualizations to quickly sort through your endpoints.

In this post, we’ll explore how NPM helps you:

To understand the flow of traffic through your system, you need to view network metrics that show you both where that traffic originated and where it terminated. The new NPM search experience allows you to combine both client and server data in one search bar for complete network visibility and streamlined troubleshooting. You can select which type of endpoints you want to filter the data for: client, server, or a combination of the two. Then you can use features like autocomplete, recent searches, and Saved Views to create useful filter queries more easily. You can also enter the tags you want to group the data by, such as cluster, host, and region. NPM displays the results broken down by the chosen client and server groupings.

A query that searches across client and server data in the NPM universal search bar.

The universal search bar can help you not only quickly find the data you’re looking for but also understand what data might be missing. For example, let’s say you’re troubleshooting a recent outage involving one of your services. After filtering your client and server communications based on the relevant service tags, you discover that data isn’t being transmitted between two key hosts. Upon further investigation, you determine that this dependency was removed during a recent update and decide to contact the appropriate team to come up with a solution.

Use clustered network maps to assess high-cardinality traffic

The NPM network map already helps you analyze traffic within your network by enabling you to trace the flow of requests and responses between your endpoints. By following these paths, you can easily identify dependencies and bottlenecks in your system.

To streamline the visualization experience for customers with a large number of endpoints, the network map comes with a cluster view that groups your endpoints automatically or by the attribute of your choice. You can sort your metrics by filters such as zone and environment, helping you to scope your data to the factors that are most relevant to your investigations. To dig deeper into a cluster, you can select it to access a high-level summary of performance metrics—including statistics on your volume, TCP, and connection data—for all the endpoints in that group.

A network map clusterd by datacenter.

You can also easily pivot to additional information about the relevant endpoints in a cluster. Let’s say that you’re troubleshooting an increase in latency within a certain region. After clustering your map based on the datacenter tag, you can pinpoint the cluster that is experiencing the issue. From here, you can jump to a view of the related hosts in NPM. The color-coded infrastructure map then helps you identify the problematic host.

Find the data you need faster with Network Performance Monitoring

Datadog Network Performance Monitoring enables you to quickly assess your network activity, giving you real-time insights that you can use to identify root causes. With NPM’s new search experience, you can streamline your network investigations by instantly locating critical metrics and inspecting your infrastructure’s dependencies. You can then visualize those metrics in the network map view to home in on unusual activity, even for very complex networks.

You can use our documentation to get started with NPM searching and clustered network maps. Or, if you’re not yet a Datadog user, 14-day free trial today.