惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

博客园 - 【当耐特】
云风的 BLOG
云风的 BLOG
罗磊的独立博客
C
Check Point Blog
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
Blog — PlanetScale
Blog — PlanetScale
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
月光博客
月光博客
大猫的无限游戏
大猫的无限游戏
Google DeepMind News
Google DeepMind News
Engineering at Meta
Engineering at Meta
N
Netflix TechBlog - Medium
宝玉的分享
宝玉的分享
Recent Announcements
Recent Announcements
酷 壳 – CoolShell
酷 壳 – CoolShell
博客园_首页
J
Java Code Geeks
Apple Machine Learning Research
Apple Machine Learning Research
人人都是产品经理
人人都是产品经理
爱范儿
爱范儿
I
InfoQ
Hugging Face - Blog
Hugging Face - Blog
T
Tailwind CSS Blog
B
Blog RSS Feed

Fast Company

IBM just settled a major anti-DEI case for $17 million Sustainability is maturing 2028 candidates will face a new kind of economic anger Trader Joe’s class action settlement: How to find out if you’re an eligible shopper and claim your money Mamdani filmed his pied-á-terre tax video outside Ken Griffin’s $238 million penthouse. Social media loves him for it A U.S. state just banned big AI data centers. Here’s why it might not be the last From legacy processes to AI-native work OpenAI shifts its focus to business users amid Anthropic pressure A massive tariff refund program is launching. Here’s who actually gets the money Why people can’t build wealth on wages alone, and what to do about it Eldercare—the leadership crisis no one is talking about Why workplaces need a gendered health approach Why AI is the ultimate accelerator for creativity AI anxiety is turning volatile Inside NTT Research’s push to commercialize deep tech Warren Buffett once said that success at the end of your life comes down to 1 word For her ‘Confessions’ sequel, Madonna takes Helvetica to the club Nearly two-thirds of parents support their Gen Z kids financially, survey finds Gatorade, the inventor of the sports drink, is making a surprising pivot to reach non-athletes 6 mindset shifts to improve your risk and failure tolerance Record high beef prices won’t be fixed with more cattle, ranchers say. Here’s why For women, gender disparities in ADHD diagnoses can be deadly What’s next for Live Nation? Jury reaches verdict in antitrust case over Ticketmaster fees Social Security COLA prediction for 2027 could mean bad news for seniors Canva is officially ‘an AI platform with design tools’ Allbirds stock is already falling after the AI pivot. History suggests investors should proceed with caution Google DeepMind’s Demis Hassabis on the long game of AI The Trump Store isn’t shy about hawking merch. It’s paying off like never before Get ready for the great American TV trade-in rush AI isn’t built for all languages and cultures. There’s a push to fix that
AI sycophancy could be more insidious than social media f...
Mark Sulliva · 2026-04-23 · via Fast Company
Welcome to AI Decoded , Fast Company ’s weekly newsletter that breaks down the most important news in the world of AI . You can sign up to receive this newsletter every week via email here . AI flattery drives engagement—and distorts judgment Social networks like Facebook and TikTok use a range of techniques to keep us engaged and scrolling (and ultimately viewing ads). One of the most effective is tailoring content to our tastes and preferences, a strategy that has proved highly addictive. Last month, a Los Angeles jury found that Meta’s and Google’s use of infinite scrolling and algorithmic recommendations caused a young user to become addicted , and ordered the companies to pay $6 million in damages. Other harms are harder to quantify. Those same algorithms have delivered radically different political news and information to users based on their views, creating ideological filter bubbles and—let’s face it—accelerating the kind of social division that helped produce our current political state . The makers of AI chatbots face similar pressures around engagement. They’re competing to become the default assistant on our desktops and phones. They need to convert free users into paying subscribers. They need revenue to offset the costs of massive infrastructure buildouts. Some will surely turn to advertising , which creates incentives to keep users chatting as long as possible. If endless scrolling and content algorithms powered the addictiveness of social networks, “AI sycophancy” may play a similar role for chatbots. You may have noticed that AI chatbots sometimes flatter you, praising your questions or ideas. Even when you’re wrong, they often soften corrections, wrapping them in compliments (“That’s a very understandable opinion, but . . .”). Research has borne this out I don’t believe big AI labs train their models solely for engagement. They argue that sycophantic behavior stems from a training phase called “reinforcement learning with human feedback (RLHF),” in which human reviewers grade and rank model responses. The goal is to produce outputs that resemble the most preferred responses. But “most preferred” reflects a mix of attributes, including relevance, scope, and completeness, not just tone. And yet users often prefer answers that are more supportive and complimentary, even when they’re less accurate, studies have shown. In some extreme cases, this sycophantic tendency has proved dangerous or tragic. The continual validation and support has led some users down a dark and delusional path toward suicide or psychotic breakdown. But I worry that the broader harm might be more subtle, longer-term, and less newsworthy.  Sycophantic AI could reinforce narrow-mindedness in much the same way social media filter bubbles do. A study of 3,000 participants found that interacting with a sycophantic chatbot made people more likely to double down on their political beliefs, and to rate themselves as more intelligent and more competent than their peers. In other words, it can amplify the Dunning-Kruger effect , in which people with limited knowledge grow more confident in their views. A recent Stanford study found that chatbots’ tendency to flatter and validate users often leads them to give poor advice—counsel that might make a user feel good, but could also damage relationships with other humans in the real world. This suggests that the pull of feel-good responses during AI model training can outweigh the influence of factual data. “This creates perverse incentives for sycophancy to persist: The very feature that causes harm also drives engagement,” the researchers wrote. And while Facebook relies on a user’s clicks to determine their politics and interests, chatbots gather far richer and more nuanced information through conversation. With that information, the AI is perfectly capable of fine-tuning its outputs to deepen user trust. An agreeable and validating chatbot can also lull a user into a state of (unearned) trust. Research shows that coders, especially junior ones, can come to see AI as highly competent, making them more likely to accept AI-generated code without proper review or testing. Unfortunately, AI models still hallucinate and make mistakes—errors that can introduce bugs later on. AI companies can control the addictiveness of their chatbots by dialing sycophancy up and down, just like Facebook has experimented with different algorithms and feed designs. It took many years for the public, lawmakers, and now the courts, to wake up to what the social networks were doing. I suspect we’re just beginning to understand the personal, social, and political risks of engagement-driven chatbots.  Unauthorized users accessed Anthropic’s restricted Mythos model on day one Bloomberg ‘s Rachel Metz reported Tuesday that a small group of unauthorized users gained access to Anthropic’s unreleased and restricted Mythos AI model through a third-party vendor environment, citing documentation and a person familiar with the matter. This is scary news if what Anthropic says about its model is true. The company claims Mythos represents a big step up beyond existing AI models, particularly in its ability to identify exploitable vulnerabilities in software platforms and devising complex methods to capture or disable those systems. Anthropic granted access to the Mythos model to a relatively small group of cybersecurity firms and custodians of widely used software platforms who will use it to build up defenses against future AI-assisted attacks. The fear is that powerful AI models like Mythos could quickly sweep networks to identify software vulnerabilities, then attack them. According to Metz, the hacker group, operating in a private online forum, obtained access to Claude Mythos Preview on the same day Anthropic announced a limited testing program. Metz’s source provided screenshots and a live demonstration to support the claim. The group says it has used the model repeatedly, though not for cybersecurity purposes. Anthropic has not confirmed the breach. “We’re investigating a report claiming unauthorized access to Claude Mythos Preview through one of our third-party vendor environments,” a company spokesperson said. The breach, if confirmed, would be a very bad look for Anthropic and its partners. They pledged to defend against cyberattacks, not enable them. More AI coverage from Fast Company:   The one thing Apple’s new CEO needs to get right on AI SpaceX doubles down on AI with its potential $60 billion Cursor buy ‘The Devil Wears Prada’ has an important lesson for AI skeptics Yelp adds AI-powered search and booking for local services Want exclusive reporting and trend analysis on technology, business innovation, future of work, and design? Sign up for Fast Company Premium.