惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

MongoDB | Blog
MongoDB | Blog
Recorded Future
Recorded Future
Jina AI
Jina AI
The Register - Security
The Register - Security
Threat Intelligence Blog | Flashpoint
Threat Intelligence Blog | Flashpoint
月光博客
月光博客
博客园 - 三生石上(FineUI控件)
F
Fortinet All Blogs
人人都是产品经理
人人都是产品经理
S
SegmentFault 最新的问题
Apple Machine Learning Research
Apple Machine Learning Research
L
LangChain Blog
Y
Y Combinator Blog
H
Hackread – Cybersecurity News, Data Breaches, AI and More
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
GbyAI
GbyAI
The GitHub Blog
The GitHub Blog
Vercel News
Vercel News
博客园 - 【当耐特】
雷峰网
雷峰网
The Cloudflare Blog
阮一峰的网络日志
阮一峰的网络日志
aimingoo的专栏
aimingoo的专栏
云风的 BLOG
云风的 BLOG
I
InfoQ
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
Google DeepMind News
Google DeepMind News
Security Latest
Security Latest
有赞技术团队
有赞技术团队
L
Lohrmann on Cybersecurity
P
Proofpoint News Feed
cs.CV updates on arXiv.org
cs.CV updates on arXiv.org
The Last Watchdog
The Last Watchdog
P
Privacy & Cybersecurity Law Blog
Scott Helme
Scott Helme
Google Online Security Blog
Google Online Security Blog
WordPress大学
WordPress大学
Hacker News - Newest:
Hacker News - Newest: "LLM"
NISL@THU
NISL@THU
cs.CL updates on arXiv.org
cs.CL updates on arXiv.org
B
Blog RSS Feed
Cyberwarzone
Cyberwarzone
K
Kaspersky official blog
F
Full Disclosure
Martin Fowler
Martin Fowler
Spread Privacy
Spread Privacy
D
Docker
C
Cisco Blogs
www.infosecurity-magazine.com
www.infosecurity-magazine.com
H
Hacker News: Front Page

Amplitude

Beyond the Rate: Retail Banking's New Competitive Front How NS Prevented €1.8M in Revenue Loss Through Experimentation Go from Product Launch to Insight to Action in Minutes What Makes a Good vs Bad North Star Metric The Role of Feature Management in Successful Product Development Cohort Retention Analysis: Reduce Churn Using Customer Data 7 Steps to Measuring the Success of a Feature 14 Best Product Management Tools for 2026 (Plus Tips from Senior PMs) The Definitive Guide to Behavioral Cohorting Putting A Number On AI Quality Meet the Winners of the 2026 Amplitude AI Impact Awards Beyond Last-Touch Attribution: Find Out Which Interactions Really Matter Agent Connectors Are Better Together Agents That Act on What Actually Happened How Square Used Amplitude to Enhance the Seller Experience and Power Growth Migrating Analytics Platforms Without The Chaos Wanted Lab Grows Sign-Ups by 150% & Builds Experimentation Culture How to Balance Inference Cost and User Experience for Agents Introducing Zoning Insights: Web Intelligence at a Glance Five best practices for getting started with AI agents 24 Quarters at #1. Here’s What’s Next. How We Built a Product That Tells Us What To Build Next: Inside Amplitude Wave Looking Beyond Campaign Metrics: 7 Marketing Success Stories AI Evals for Product Managers: A Beginner’s Guide to Getting Started The Builder Skills Library Introducing Agent Connectors in Amplitude Understand How AI Thinks, Get Better Results How We Redesigned Amplitude Docs for Agents and Made Everyone an Author AI Broke Your Experimentation Program. Here’s How to Fix It. Every Stuck User Is a Support Ticket Waiting to Happen Tracing the Sale: Connect Behavior to Conversions with Persisted Properties Building CLI Agents: It’s What You Don’t Give Them That Counts Three Tips for Better Prompts in Amplitude Global Agent How AI Took the Data Analyst’s Job, and Created a Better One Default Prompts Are Tanking Your Agent’s Retention Optimizing Core Web Vitals with Amplitude’s Global Agent Don’t Ask Global Agent Anything, Ask These Three Things How We Built a Design Agent at Amplitude with Claude Managed Agents and Cloudflare The Problem with Chasing Churn How Hostinger Achieved a 20%+ Conversion Lift Through Experimentation How STAGE Streams Smarter by Putting Data at the Center Building the Validation Stack for AI Product Development Making AI Analytics Safe for Financial Services Teams Amplitude Heatmaps Update: More Reliable Screenshots and Accurate Placement Most Teams Ship Agent Personalities by Accident. We Didn’t. What I Learned Pointing a Ralph Loop at My Product for a Week How Mercado Libre Scales Decision Making with AI Claude Cowork for PMs: 5 Playbooks to Get Started How ACKO Drove 13% More Conversions & 50% Drop in Calls with GenAI Agents Just Made Your Feature Launch Channel Smarter Homegrown FinOps Tools: How AI “Build” Beat “Buy” for Us in <1 Year Introducing The Amplitude Quickstart Series Rebuilding Session Replay’s Delivery Layer to Be Lighter on Your Page The Eval Signal That Predicts 3x Agent Retention Amplitude and Statsig Partnership 5 Agent Skills to Automate Your Weekly Product Review Amplitude Plug and Play: New AI Plugin in Claude and Cursor Marketplaces Introducing Amplitude Wizard CLI: Set Up Amplitude from Your Codebase Making AI Search Count (and Convert) How VEED Evolved Its AI Search Strategy What’s New with Amplitude Agents Effortless Support at Scale: Making Human Support More Human AI Week 2026: Upleveling All Together Amplitude AI Builders: Paul Hultgren Chats about AI Assistant Dashboard Dread to AI-Driven Decisions: How Tira Rebuilt Its Analytics Workflow Your Product Deserves a Better Support Agent How Cisco Systems Accelerated Adoption by 20% Through Data Innovation
Agents Write Code. Fixing It Is Still On You.
Chanaka Perera · 2026-05-06 · via Amplitude

This blog was co-authored by Eric Kim, Head of Engineering, Agents at Amplitude.

Agents are writing more code than ever, but when something breaks in production, the investigation looks the same. You’re pulled away from the feature you’re shipping to investigate the bug report in Linear, check logs in Datadog, and comb your session replay tool to figure out what went wrong.

Amplitude MCP brings all of that session data directly into Claude and Cursor, so the investigation happens in the same place where your agent will write the code. Now the bugs you used to skip become ones you can actually fix.

Investigate bugs in real time

You get an urgent bug report in Jira or Linear and it’s time to investigate the fix. With Amplitude MCP, you can call Session Replay directly in Claude or Cursor. Here’s what it looks like.

First, describe the bug in plain language (e.g., “Users are having trouble checking out, what’s going on?”). The right skill triggers automatically based on what you ask:

  • If you already have a concrete starting point, like a user report or a specific error name, the debug-replay skill can reproduce it.
  • If you only have a vague issue, like “the checkout flow doesn’t work,” then diagnose-errors can figure out what’s broken.
  • If you want a reliability check across sessions, monitor-reliability will trigger.

If your team has instrumented events to monitor for the specific errors and issues the bug is related to, then these skills will orchestrate a workflow, so you don’t have to go through each step manually.

If you’re still having trouble reproducing the bug, or if you want more control over the investigation, try these tips to narrow down the cause:

  • Find the sessions where the bug happened. get_session_replays retrieves candidate sessions matching an error, user, time window, or event, including specific error events your app has instrumented.
  • See what the user actually did. get_session_replay_events extracts the full interaction timeline, including every click, event, and console error.
  • Correlate with deployments. get_deployments checks whether the bug aligns with a recent release.

Diagnostic information is helpful, but visualizing the bug can help you validate and add more detail to your investigation. Ask your agent to “Find the session where the bug happened,” and narrow it down by user email, time, and date. The replay will render directly in Claude or Cursor, confirming what’s broken so the agent can write the fix.

Investigating and reproducing a bug used to be the slow part of fixing it. Now, what took half a day of context switching happens in a single session in a single tool.

Your bug investigation cheat sheet

Catch friction before it becomes a ticket

Not every bug starts as a Linear ticket or an urgent Slack message. Sometimes, the bug never shows up at all, and users leave without saying a word. Proactively spotting these instances of friction and failure protects your users from frustration and churn.

Amplitude’s session replay agent runs in the background to continuously watch user sessions and surface these patterns before they show up in your queue. It regularly reviews sessions, flags friction signals, and posts a weekly summary to Slack.

When you notice a new friction pattern emerging, you can pull the agent’s report directly into Claude or Cursor to investigate and fix the issue. Use get_agent_results to return the agent’s analysis: a narrative summary of the friction, the pattern type, representative session IDs, impact framing, and recommended next steps. Now you’re no longer starting from a blank page.

Next, validate the pattern with actual sessions. Use get_session_replay_events to pull the events and see the interaction timeline, or ask your agent to find and render the relevant replays directly in Claude or Cursor.

Once you’ve validated the issue, decide if it needs a fix now. Some issues are worth pulling into your backlog but don’t need a same-day fix. And if it’s urgent, the agent already has the context loaded to ship the fix.

Use this workflow to get ahead of issues before they become a fire drill or lead to invisible customer churn.

Fix the bugs you used to defer

Debugging used to mean leaving your code to investigate: pulling data logs, finding the right session, and scrubbing through replays. The investigation often took longer than the fix itself.

These workflows fix that. The urgent bug lands in your inbox, and the investigation happens in the same place where your agent writes the code. The friction pattern surfaces in Slack, and you pull the agent’s analysis straight into Cursor or Claude.

The point of these workflows isn’t just faster debugging. It changes what bugs get fixed at all. If investigation takes an hour, you’re only ever prioritizing the highest tickets in your queue. If it takes ten minutes, you can work through a class of bugs you used to always defer.

Agents write your code. With these workflows, they can help fix it too.