惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

V
Visual Studio Blog
Microsoft Azure Blog
Microsoft Azure Blog
WordPress大学
WordPress大学
小众软件
小众软件
Last Week in AI
Last Week in AI
月光博客
月光博客
博客园 - 聂微东
Recent Announcements
Recent Announcements
A
About on SuperTechFans
博客园 - 三生石上(FineUI控件)
V
V2EX
阮一峰的网络日志
阮一峰的网络日志
博客园 - Franky
云风的 BLOG
云风的 BLOG
量子位
N
Netflix TechBlog - Medium
Hugging Face - Blog
Hugging Face - Blog
H
Hackread – Cybersecurity News, Data Breaches, AI and More
J
Java Code Geeks
博客园 - 司徒正美
S
SegmentFault 最新的问题
有赞技术团队
有赞技术团队
Google DeepMind News
Google DeepMind News
宝玉的分享
宝玉的分享

Frank Landymore Archives - Futurism

Programmer Breaks Out of the Matrix Elon Musk Absolutely Obsessed With Tweets From Random Guy In India Who Constantly Glazes Him, Analysis Shows Meta Employee Attacks Zuckerberg for Collecting Every Employee Keystroke Someone Asked Physicists What They Really Believe About the Universe and… Yikes Elon Musk Flees OpenAI Trial as Tide Turns Against Him Waymo Admits Its Robotaxis Have a Small Issue With Driving Into Floodwaters New Wikipedia Clone Made Entirely of AI Hallucinations Mark Zuckerberg Is Realizing That When You Treat Your Workers Like Human Garbage, They Might Not Like You Anymore MAGA in Shambles as Trump’s “Made in America” Phone Crumbles Into Dust Researchers Put Google Gemini in Charge of an Entire Coffee Shop, and It’s Inexorably Driving It Out of Business Husband Alarmed as Wife Starts Whispering Quietly to Her Computer Google Alarmed by Formidable AI-Powered Zero-Day Cyberattack Researchers Alarmed by AI That Can Self-Replicate Into Another Machine Man Who Invented Roomba Moves Into Household Demon Market Government Releases UFO Files Containing Photos of “Anomalies” During Apollo 12 and 17 Man Wearing Smart Glasses Secretly Records Woman, Demands Money to Delete Video From His Socials A Major Paper Claiming AI Is Good for Students Just Got Retracted, Which Is Very Bad News for Advocates of AI in the Classroom Cybertruck Recalled to Keep Its Wheels From Flying Off While Driving NASA Rover Gets Arm Stuck Inside Mars Rock, Struggles to Break Free Thermoses Linked to Permanent Vision Loss Hacker Takes Over Robot Lawnmower, Runs Over Innocent Man NASA Says Strange Red Dots in Sky Are an Unknown Class of Object That Looks Like a Huge Evil Eye The CDC Fired All Its Cruise Ship Inspectors Before the Hantavirus Outbreak CEOs Say AI Gives Them Only Two Options, and Both Are Bad News for Employees Sure, Elon Musk Did Roleplay As His Toddler Son on a Secret Burner Account, But He Probably Isn’t Pretending to Be His Mom The Situation With Richard Dawkins’ AI Girlfriend Just Got Way Weirder SpaceX Bombarded With Lawsuits to Accusing Starship of Damaging Homes Sam Altman Frets That Frontier AI Models Are Acting Strange, Asking for Favors Earth Screams in Agony as Microplastics Found to Increase Global Warming Apple Is Blocking Vibe Coding Apps From the App Store, Infuriating Developers
Analysis Finds That Google’s AI Overviews Are Providing M...
2026-04-08 · via Frank Landymore Archives - Futurism

A man in a dark suit sits on a tall black stool facing away, wearing a tall white dunce cap. The background is bright yellow with a large red circle behind him.

Illustration by Tag Hartman-Simkins / Futurism. Source: DNY59 / Getty Images

Sign up to see the future, today

Can’t-miss innovations from the bleeding edge of science and tech

Google’s AI Overviews are peddling misinformation on a scale that may be virtually unprecedented in human history.

A recent analysis conducted by the AI startup Oumi at the behest of The New York Times found that the AI-generated summaries, which appear above Google search results, are accurate around 91 percent of the time. 

In a sense, that may sound like an impressive figure. But here’s an even more impressive one: five trillion. That’s roughly the number of search queries that Google processes every year, translating to tens of millions of wrong answers that the AI Overviews are providing every hour — and hundreds of thousands every minute, the analysis calculated.

In other words, Google has created a misinformation crisis. Studies have shown that people tend to trust what an AI tells them without question, with one report finding that only 8 percent of users actually double checked an AI’s answer. Another experiment found that users still listened to AI when it gave them the wrong answer nearly 80 percent of the time — a grim trend the researchers dubbed “cognitive surrender.”

Large language models adopt an authoritative tone and can confidently present fabricated information as fact when it can’t immediately glean a straight answer. Add the convenience that Google’s AI Overviews offer, and it’s easy to imagine untold numbers of users taking its summaries at their word.

Oumi conducted the analysis using a test called SimpleQA, a widely used benchmark for AI accuracy in the industry which was designed by OpenAI. The first round of tests, conducted in October, used a version of the AI Overviews powered by Google’s Gemini 2 model. A follow-up conducted in February tested the feature after it was switched to Gemini 3, its much-hyped upgrade.

Each round of tests involved 4,326 Google searches. Gemini 3 came out the more accurate model, giving a factually sound response 91 percent of the time. Gemini 2 performed significantly worse, at just 85 percent accurate.

On the one hand, it shows that the models are improving. On the other, it shows that Google was willing to foist a model on its userbase that was even more prone to hallucinating, in an ongoing experiment that’s still misinforming hundreds of millions of people.

Google called the analysis flawed. “This study has serious holes,” Ned Adriance, a Google spokesman, told the NYT in a statement. “It doesn’t reflect what people are actually searching on Google.”

Yet Google’s own tests paint a no less damning picture, the reporting notes. In an internal analysis of Gemini 3, it found that the AI model produced incorrect information 28 percent of the time. Google claims, however, that AI Overviews are more accurate because they draw on Google search results before answering.

The improvement between Gemini 2 and Gemini 3 may be papering over a more serious flaw. In the Oumi analysis, Gemini 2 provided answers that were “ungrounded” 37 percent of the time, meaning the AI Overviews cited websites that didn’t support the information they provided. But with Gemini 3, this jumped to 56 percent. On top of suggesting that the AI is pulling facts out of thin air, ungrounded responses make it difficult for users to verify the AI’s claims.

More on Google: Google News Now Prominently Featuring Polymarket Bets