惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

G
Google Developers Blog
GbyAI
GbyAI
Y
Y Combinator Blog
The GitHub Blog
The GitHub Blog
B
Blog
博客园 - 叶小钗
V
Visual Studio Blog
小众软件
小众软件
阮一峰的网络日志
阮一峰的网络日志
博客园 - 聂微东
S
SegmentFault 最新的问题
Engineering at Meta
Engineering at Meta
博客园 - Franky
V
V2EX
人人都是产品经理
人人都是产品经理
H
Hackread – Cybersecurity News, Data Breaches, AI and More
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
月光博客
月光博客
IT之家
IT之家
T
The Blog of Author Tim Ferriss
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
C
Check Point Blog
N
Netflix TechBlog - Medium
博客园 - 【当耐特】

Datadog | The Monitor blog

Introducing our open source AI-native SAST Instrument and monitor Boomi integration flows with OpenTelemetry and Datadog Not all index scans are equal: How we cut query latency by over 99% Platform engineering metrics: What to measure and what to ignore Integrate Recorded Future threat intelligence with Datadog Cloud SIEM CI/CD security: threat modeling using a MITRE-style threat matrix CI/CD security: How to secure your GitHub ecosystem Ingress NGINX is EOL: A practical guide for migrating to Kubernetes Gateway API Operating agentic AI with Amazon Bedrock AgentCore and Datadog LLM Observability: Lessons from NTT DATA Introducing the Datadog Code Security MCP Capture and analyze custom heatmaps in Session Replay Understand session replays faster with AI summaries and smart chapters Monitor ClickHouse query performance with Datadog Database Monitoring How we designed empathetic alert sounds for on-call engineers Search and act across Datadog to resolve issues faster with Bits Assistant Measure the business impact of every product change with Datadog Experiments Analyzing round trip query latency Configuring JavaScript caches for better performance Introducing Bits AI Dev Agent for Code Security Datadog achieves ISO 42001 certification for responsible AI Monitor Nutanix clusters, hosts, and VMs with Datadog Monitor Juniper Mist in Datadog A new Host Map for modern infrastructure Annotate traces to improve LLM quality with Datadog LLM Observability What’s new in Cloud SIEM: AI-powered investigations, enhanced threat intelligence, and scalable security operations Explore Kubernetes with native OpenTelemetry data Monitor Oracle Fusion Cloud Applications with Datadog Announcing the Datadog Terraform provider v4.0.0 Scaling Kubernetes workloads on custom metrics How to design cloud environments for AI-powered threat analysis
Streamline network investigations with an enhanced queryi...
2023-05-08 · via Datadog | The Monitor blog

Editor’s note: This post covers Cloud Network Monitoring, a Datadog feature that was originally called Network Performance Monitoring.

Effective network troubleshooting requires collecting and correlating thousands of data points across your entire stack. The more data you ingest, however, the more data you have to search through in order to locate important signals. This can make it hard to find the information you need during time-sensitive investigations.

Datadog Network Performance Monitoring (NPM) provides enhanced search and mapping capabilities that help you quickly pinpoint critical metrics, giving you clear visibility into the health of your network at any given time. NPM’s new search experience delivers streamlined, unified filtering that enables you to comb through both client and server data simultaneously for a complete picture of communication across your network. And with clustered maps, you can organize large quantities of network data into neatly categorized visualizations to quickly sort through your endpoints.

In this post, we’ll explore how NPM helps you:

To understand the flow of traffic through your system, you need to view network metrics that show you both where that traffic originated and where it terminated. The new NPM search experience allows you to combine both client and server data in one search bar for complete network visibility and streamlined troubleshooting. You can select which type of endpoints you want to filter the data for: client, server, or a combination of the two. Then you can use features like autocomplete, recent searches, and Saved Views to create useful filter queries more easily. You can also enter the tags you want to group the data by, such as cluster, host, and region. NPM displays the results broken down by the chosen client and server groupings.

A query that searches across client and server data in the NPM universal search bar.

The universal search bar can help you not only quickly find the data you’re looking for but also understand what data might be missing. For example, let’s say you’re troubleshooting a recent outage involving one of your services. After filtering your client and server communications based on the relevant service tags, you discover that data isn’t being transmitted between two key hosts. Upon further investigation, you determine that this dependency was removed during a recent update and decide to contact the appropriate team to come up with a solution.

Use clustered network maps to assess high-cardinality traffic

The NPM network map already helps you analyze traffic within your network by enabling you to trace the flow of requests and responses between your endpoints. By following these paths, you can easily identify dependencies and bottlenecks in your system.

To streamline the visualization experience for customers with a large number of endpoints, the network map comes with a cluster view that groups your endpoints automatically or by the attribute of your choice. You can sort your metrics by filters such as zone and environment, helping you to scope your data to the factors that are most relevant to your investigations. To dig deeper into a cluster, you can select it to access a high-level summary of performance metrics—including statistics on your volume, TCP, and connection data—for all the endpoints in that group.

A network map clusterd by datacenter.

You can also easily pivot to additional information about the relevant endpoints in a cluster. Let’s say that you’re troubleshooting an increase in latency within a certain region. After clustering your map based on the datacenter tag, you can pinpoint the cluster that is experiencing the issue. From here, you can jump to a view of the related hosts in NPM. The color-coded infrastructure map then helps you identify the problematic host.

Find the data you need faster with Network Performance Monitoring

Datadog Network Performance Monitoring enables you to quickly assess your network activity, giving you real-time insights that you can use to identify root causes. With NPM’s new search experience, you can streamline your network investigations by instantly locating critical metrics and inspecting your infrastructure’s dependencies. You can then visualize those metrics in the network map view to home in on unusual activity, even for very complex networks.

You can use our documentation to get started with NPM searching and clustered network maps. Or, if you’re not yet a Datadog user, 14-day free trial today.