惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Apple Machine Learning Research
Apple Machine Learning Research
爱范儿
爱范儿
博客园_首页
博客园 - 【当耐特】
V
Visual Studio Blog
博客园 - 叶小钗
月光博客
月光博客
美团技术团队
J
Java Code Geeks
小众软件
小众软件
Y
Y Combinator Blog
博客园 - Franky
Martin Fowler
Martin Fowler
博客园 - 聂微东
Microsoft Azure Blog
Microsoft Azure Blog
IT之家
IT之家
MyScale Blog
MyScale Blog
人人都是产品经理
人人都是产品经理
Microsoft Security Blog
Microsoft Security Blog
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
阮一峰的网络日志
阮一峰的网络日志
酷 壳 – CoolShell
酷 壳 – CoolShell
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
云风的 BLOG
云风的 BLOG

Datadog | The Monitor blog

Introducing our open source AI-native SAST Instrument and monitor Boomi integration flows with OpenTelemetry and Datadog Not all index scans are equal: How we cut query latency by over 99% Platform engineering metrics: What to measure and what to ignore Integrate Recorded Future threat intelligence with Datadog Cloud SIEM CI/CD security: threat modeling using a MITRE-style threat matrix CI/CD security: How to secure your GitHub ecosystem Ingress NGINX is EOL: A practical guide for migrating to Kubernetes Gateway API Operating agentic AI with Amazon Bedrock AgentCore and Datadog LLM Observability: Lessons from NTT DATA Introducing the Datadog Code Security MCP Capture and analyze custom heatmaps in Session Replay Understand session replays faster with AI summaries and smart chapters Monitor ClickHouse query performance with Datadog Database Monitoring How we designed empathetic alert sounds for on-call engineers Search and act across Datadog to resolve issues faster with Bits Assistant Measure the business impact of every product change with Datadog Experiments Analyzing round trip query latency Configuring JavaScript caches for better performance Introducing Bits AI Dev Agent for Code Security Datadog achieves ISO 42001 certification for responsible AI Monitor Nutanix clusters, hosts, and VMs with Datadog Monitor Juniper Mist in Datadog A new Host Map for modern infrastructure Annotate traces to improve LLM quality with Datadog LLM Observability What’s new in Cloud SIEM: AI-powered investigations, enhanced threat intelligence, and scalable security operations Explore Kubernetes with native OpenTelemetry data Monitor Oracle Fusion Cloud Applications with Datadog Announcing the Datadog Terraform provider v4.0.0 Scaling Kubernetes workloads on custom metrics How to design cloud environments for AI-powered threat analysis
Data-driven storytelling with Datadog Notebooks
Abril Loya McCloud · 2017-01-10 · via Datadog | The Monitor blog
Abril Loya McCloud

Abril Loya McCloud

When an incident disrupts availability or performance, you want to be able to investigate, correct the problem, identify warning signs for the future, and document it all. Usually, this involves seeking out information from different services, gathering metrics and screenshots, then compiling and distributing your findings through email, wikis, or text docs.

Datadog’s new Notebooks feature allows you to combine real-time or historical graphs with Markdown cells to:

Better postmortems

Notebooks allow you to create detailed postmortems that you can share with your entire team. Building around graphs from the incident, you can add text cells to explain and contextualize an incident, its cause, and what was done to remedy the situation. Because Notebooks support Markdown, you can easily organize your postmortem using headers, or add formatting like lists and code snippets.

Data-driven postmortems

Every graph in a notebook can be set to an adjustable “Global Time” or locked to its own specific timeframe. So you can show graphs that depict system behavior at the time of the incident as well as metrics leading up to the event. Pinpointing system behavior leading up to the incident provides the information you need to create alerts that can help you get ahead of the issue next time.

Graph with adjustable time

Dynamic runbooks

Runbooks help members of your team respond to issues by providing them with detailed instructions and historical context. For runbooks to be helpful, team members need them to be accessible and up to date. Because Notebooks are accessible to anyone in your Datadog organization, they make it easy to distribute and collaboratively update runbooks.

When an alert is triggered, having a runbook can make all the difference in response time. By providing a link to the relevant runbook in your Datadog alerts, you can ensure that whoever is on-call receives step-by-step instructions for dealing with known issues.

Data-driven runbooks

Open-ended exploration

With Notebooks, you can quickly explore any of your infrastructure or application metrics. Metrics can be visualized as timeseries, heatmaps, or distributions. You can also compare metric performance across groups—for instance, you can break out a graph of a globally aggregated metric into individual graphs for each availability zone.

Data-driven exploration

New Notebooks are unsaved by default, so you can visualize your metrics without worrying about modifying your existing dashboards or cluttering up your list of production dashboards with one-off scratch pads. If you discover something worth saving or sharing, however, you can save your work with the “Save Notebook” button.

Go forth and explore!

Rich, clear, easily accessible internal documentation provides much-needed context to engineering teams. The new Notebooks feature in Datadog makes it easy to create and maintain postmortems and runbooks, while also allowing you to explore your metrics freely.

If you’re already a Datadog customer, you can access Notebooks by clicking on the “Notebooks” button in your sidebar. Otherwise, you can sign up for a free 14-day trial and introduce data-driven storytelling to your organization today.