惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

博客园_首页
H
Help Net Security
腾讯CDC
宝玉的分享
宝玉的分享
H
Hackread – Cybersecurity News, Data Breaches, AI and More
L
LangChain Blog
爱范儿
爱范儿
T
The Blog of Author Tim Ferriss
J
Java Code Geeks
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
MyScale Blog
MyScale Blog
Engineering at Meta
Engineering at Meta
N
Netflix TechBlog - Medium
D
Docker
V
V2EX
Last Week in AI
Last Week in AI
G
Google Developers Blog
IT之家
IT之家
C
Check Point Blog
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
人人都是产品经理
人人都是产品经理
博客园 - 叶小钗
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
博客园 - 聂微东

Forbes - Innovation

Why Do Humans Have Fingerprints? Hint: It’s Not What You Think Booking.com Confirms Data Breach, Reservation PIN Codes Changed Why Major News Sites Are Blocking The Internet Archive’s Wayback Machine iPhone Fold Release Date: New Report Details Frustrating Apple News Comet Tracker: How To See Pan-STARRS And Three Planets On Wednesday NYT Mini Crossword Today: Tuesday, April 14 Hints And Answers Today’s NYT Strands Hints, Spangram, Answers: Tuesday, April 14 (It’s A Little Unclear) Today’s Wordle #1760 Hints And Answer For Tuesday, April 14 Most Of The Microplastics In Urban Air Come From Tires Today’s Wordle #1759 Hints And Answer For Monday, April 13 NYT Mini Crossword Today: Monday, April 13 Hints And Answers NYT Pips Today: Hints, Answers And Walkthrough For Monday, April 13 The YC Chief Who Codes 10,000 Lines A Day Has A Simple Secret Samsung Expands One UI 8.5 Beta To More Galaxy Owners Why You Should Stop Using Your iPhone If It’s On This List Chamath Says Firms That Treat AI As A Strategy Hand Rivals Their Edge 3 Unexpected Habits Of Secure Couples, By A Psychologist The First Lamp That Folds Your Clothes Samsung’s Disappointing Price Update For Galaxy Phone Buyers 3 Subtle Signs Someone Is Falling In Love With You, By A Psychologist Do Mantis Shrimp See More Colors Than Humans? A Biologist Explains NYT Connections Answers Explained For Monday, April 13 (#1,037) NYT Connections Hints Today: Monday, April 13 Clues And Answers (#1,037) LEGO Luigi & Mach 8 (72050) Review: 2026’s Best Set Yet? Marc Andreessen Says AI Productivity Will Trigger A Hiring Boom 3D Printing Is The Ultimate Hack To Reduce Household Spending Apple iPhone Fold: Striking Design Revealed In Leaked Photos Apple Smart Glasses: New Leak Reveals A Major Design Twist To Beat Meta Tested: The AI Coming To The Rivian R2 Quordle Hints Today: Monday, April 13 Clues And Answers
​Your AI Agent Thinks It's Right, And That's Exactly The ...
Stu Sjouwerman · 2026-06-25 · via Forbes - Innovation

Stu Sjouwerman is co-founder and CEO of ReadingMinds, a pioneering AI-moderated interview platform for conducting sentiment analysis.

getty

One of the most powerful aspects of AI is the ability for agents to learn and, presumably, to improve over time. In the early years of agentic AI, agents operated session by session. They completed a task and then started from scratch with the next task. What they learned from previous tasks wasn’t carried over. That’s no longer the case.

For example, new persistent memory capabilities were unlocked in ChatGPT. Microsoft embedded memory straight into Copilot. All of this sounds good and valuable, and it can be. But it can also create the potential for hiccups that can diminish the value of these tools and potentially impact performance and brand value.

Interpretation Vs. Facts

Although we have a tendency to think that AI agents can actually think like us, they don’t at all. For instance, when AI agents refer back to customer interactions, they are not quoting a transcript directly. Rather, they are relying on an interpretation of what the technology thinks it recognizes. There's a crucial difference that could be dangerous.

Let me show you what this looks like in a common interaction with a customer.

A B2B software company contacts you during its early-stage research for possible product solution(s). Let us say that during an interaction, a prospect hesitates before responding to a question related to timing. The agent interprets that as low urgency, and this conclusion is logged. That stored impression is then fed into all future interactions and interpretations.

In this instance, bear in mind that the prospect might have paused because the internal budget process was shifting or had changed. They were interested, but they didn’t have a ready response to the question.

However, due to the way the agent learned from this prospect, the lead gets stuck, and the prospect goes on to find a solution elsewhere. The sale is lost.​

When Agents Act With Too Much Confidence

Here’s something that should legitimately cause concern among C-suite executives. AI agents don’t know when they make an incorrect interpretation. They act with unwavering confidence. That’s the nature of how they operate.

AI models use confidence scores to drive future actions. For instance, if the model exceeds a given threshold, the agent will proceed without human intervention. While the process is intended to support reasonable governance, in reality, it does not.

Confidence scores measure some level of internal coherence and how consistent the model’s reasoning is based on its training and stored content.

Here’s what these scores don’t measure: whether the underlying interpretation reflects reality.

This is referred to as a calibration problem. It’s the gap between the certainty that a model exhibits and how accurate the interpretation actually is.

LLMs have a documented tendency toward overconfidence, expressing high certainty even in situations that turn out to be incorrect. The impacts of this misguided certainty compound across months of customer interactions. You can see how this problem can escalate in large enterprise deployments.

A Governance Problem

This isn’t so much a technical problem as it is a governance problem, which can worsen over time. Like a cascade, a stored misinterpretation can shape the next decision, and the next, and the next. Each decision generates new data, which the agent considers to be a confirmation. The system then becomes progressively more confident. By the time the data are viewed by a human, the pattern looks deceptively coherent because every subsequent step was built on the same flawed interpretation.

Envision this same cascade occurring across hundreds of accounts that are connected to your CRM, email and pipeline management. Some single misread that occurred at the beginning of a relationship is perpetuated across every touchpoint that the agent controls.

Speed, the main hallmark that makes agentic AI so compelling, suddenly becomes a liability.

Agents don’t make singular mistakes. They make multiple mistakes rapidly at scale, and in a blinding way that appears to convey righteousness and competence.

A Takeaway For Executives

Rather than focusing on scale and how much an agent can retain, ask how the agent knows when it has learned something incorrectly. That correction cannot come from within the agent's own reasoning loop. An agent can’t reliably detect its own misinterpretations using the same model that generated them.

Validation must come from an external source. What does this mean?

It means that in practice, validation must occur at moments when customers are expressing themselves.

Perhaps during voice interviews at key points of interaction. Perhaps through well-designed signals, captured prior to agents acting on assumptions stored in memory. Perhaps through a continuous layer designed to flag when a customer’s expressed signals do not match what is coded in the system.

These signals don’t replace agent memory but correct it. Memory without correction isn’t intelligence; it’s what you call a bias. In customer-facing workflows, bias compounds.

What matters most isn’t deploying the most capable agents. It’s putting the right architecture in place to ensure that those agents remain honest and accurate. ​


Forbes Technology Council is an invitation-only community for world-class CIOs, CTOs and technology executives. Do I qualify?