惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

B
Blog
The Cloudflare Blog
J
Java Code Geeks
Apple Machine Learning Research
Apple Machine Learning Research
T
Tailwind CSS Blog
L
LangChain Blog
Recent Announcements
Recent Announcements
Hugging Face - Blog
Hugging Face - Blog
Microsoft Security Blog
Microsoft Security Blog
F
Fortinet All Blogs
Microsoft Azure Blog
Microsoft Azure Blog
V
V2EX
I
InfoQ
博客园 - 司徒正美
T
The Blog of Author Tim Ferriss
G
Google Developers Blog
云风的 BLOG
云风的 BLOG
aimingoo的专栏
aimingoo的专栏
小众软件
小众软件
H
Help Net Security
博客园 - 三生石上(FineUI控件)
S
SegmentFault 最新的问题
B
Blog RSS Feed
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知

Futurism

Meta Installing Software on Employee Computers to Track Everything They Do, Feed the Data to AI Chinese Workers Horrified as Bosses Direct Them to Train Their AI Replacements Concern Grows That AI Is Damaging Users’ Cognitive Abilities JPMorganChase Data Center Gets $77 Million Handout to Create Grand Total of One Job Nvidia CEO Loses His Cool at Tough Question CEO of $1.5 Billion AI Startup Accused of Massive Fraud by Justice Department Palantir Issues Ominous Corporate Manifesto Madison Square Garden Reportedly Used Facial Recognition to Stalk Trans Woman For Two Years The Florida Mass Shooter’s Conversations With ChatGPT Are Worse Than You Could Possibly Imagine China Is Starting to Pull Ahead of US in AI Race AI Company Known for Teen Suicides Launches New Feature to Turn Books Into Roleplaying Experiences Study Finds AI Use Eats Away at Users’ Confidence in Their Own Brains Democrats Warned Not to Upset Multi-Million Dollar AI Lobbyists, Even Though It’d Be a Slam Dunk With Voters City Council Wrecked in Voter Bloodbath After Allowing New Data Center Mother Reportedly Doesn’t Know Her Son Died Because She’s Been Talking to an AI Version of Him Things You Told ChatGPT or Claude My Have Already Doomed You in Court Millions of Americans Are Talking to AI Instead of Going to the Doctor, and It’s Giving Them Horrendously Flawed Medical Advice There Are Signs of a Massive AI Backlash A Prominent PR Firm Is Running a Fake News Site That’s Plagiarizing Original Journalism at Incredible Scale Fury Erupts as Val Kilmer’s Estate Announces Starring Role in AI Film Made From Beyond the Grave Allbirds Stock Now Crashing as Reality Sets in About Its Delusional AI Pivot NAACP Sues Elon Over His Noxious AI Data Center Top Security Experts Alarmed by Power of Anthropic’s New Hacker AI Teens Alarmed at What AI Is Doing to Their Minds What It Really Means That a Failing Shoe Brand “Pivoted to AI” and Its Stock Soared 700 Percent Starbucks’ Baffling ChatGPT Collab Treats Customers Like Empty, Soulless Venti Cups ChatGPT’s “Honest Reaction” to a “Song” Composed Entirely of Gas-Passing Noises Will Make You Question Whether It’s Honestly Evaluating Your Other Brilliant Ideas AI Is Turning Workplaces Into Hopeless Gridlock Companies Just Learned a Brutal Lesson About Training AI to Do Human Jobs Berklee College of Music Students Furious That It’s Offering an AI “Songwriting” Class
Analysis Finds That Google’s AI Overviews Are Providing M...
2026-04-08 · via Futurism

A man in a dark suit sits on a tall black stool facing away, wearing a tall white dunce cap. The background is bright yellow with a large red circle behind him.

Illustration by Tag Hartman-Simkins / Futurism. Source: DNY59 / Getty Images

Sign up to see the future, today

Can’t-miss innovations from the bleeding edge of science and tech

Google’s AI Overviews are peddling misinformation on a scale that may be virtually unprecedented in human history.

A recent analysis conducted by the AI startup Oumi at the behest of The New York Times found that the AI-generated summaries, which appear above Google search results, are accurate around 91 percent of the time. 

In a sense, that may sound like an impressive figure. But here’s an even more impressive one: five trillion. That’s roughly the number of search queries that Google processes every year, translating to tens of millions of wrong answers that the AI Overviews are providing every hour — and hundreds of thousands every minute, the analysis calculated.

In other words, Google has created a misinformation crisis. Studies have shown that people tend to trust what an AI tells them without question, with one report finding that only 8 percent of users actually double checked an AI’s answer. Another experiment found that users still listened to AI when it gave them the wrong answer nearly 80 percent of the time — a grim trend the researchers dubbed “cognitive surrender.”

Large language models adopt an authoritative tone and can confidently present fabricated information as fact when it can’t immediately glean a straight answer. Add the convenience that Google’s AI Overviews offer, and it’s easy to imagine untold numbers of users taking its summaries at their word.

Oumi conducted the analysis using a test called SimpleQA, a widely used benchmark for AI accuracy in the industry which was designed by OpenAI. The first round of tests, conducted in October, used a version of the AI Overviews powered by Google’s Gemini 2 model. A follow-up conducted in February tested the feature after it was switched to Gemini 3, its much-hyped upgrade.

Each round of tests involved 4,326 Google searches. Gemini 3 came out the more accurate model, giving a factually sound response 91 percent of the time. Gemini 2 performed significantly worse, at just 85 percent accurate.

On the one hand, it shows that the models are improving. On the other, it shows that Google was willing to foist a model on its userbase that was even more prone to hallucinating, in an ongoing experiment that’s still misinforming hundreds of millions of people.

Google called the analysis flawed. “This study has serious holes,” Ned Adriance, a Google spokesman, told the NYT in a statement. “It doesn’t reflect what people are actually searching on Google.”

Yet Google’s own tests paint a no less damning picture, the reporting notes. In an internal analysis of Gemini 3, it found that the AI model produced incorrect information 28 percent of the time. Google claims, however, that AI Overviews are more accurate because they draw on Google search results before answering.

The improvement between Gemini 2 and Gemini 3 may be papering over a more serious flaw. In the Oumi analysis, Gemini 2 provided answers that were “ungrounded” 37 percent of the time, meaning the AI Overviews cited websites that didn’t support the information they provided. But with Gemini 3, this jumped to 56 percent. On top of suggesting that the AI is pulling facts out of thin air, ungrounded responses make it difficult for users to verify the AI’s claims.

More on Google: Google News Now Prominently Featuring Polymarket Bets