惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

MyScale Blog
MyScale Blog
Jina AI
Jina AI
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
阮一峰的网络日志
阮一峰的网络日志
P
Proofpoint News Feed
Last Week in AI
Last Week in AI
博客园 - 司徒正美
Martin Fowler
Martin Fowler
T
Tailwind CSS Blog
B
Blog RSS Feed
Vercel News
Vercel News
博客园 - 聂微东
I
InfoQ
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
M
MIT News - Artificial intelligence
Recent Announcements
Recent Announcements
GbyAI
GbyAI
L
LangChain Blog
酷 壳 – CoolShell
酷 壳 – CoolShell
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
Microsoft Security Blog
Microsoft Security Blog
C
Check Point Blog
MongoDB | Blog
MongoDB | Blog
B
Blog

Futurism

Man Creates Tiny Submarine for His Parakeet to Experience Life Underwater The Effects of AI-Generated Code Tearing Through Corporations Is Actually Kind of Funny Trump Hires Orbital Towing Company to Build Space Interceptors Psychologists Found Something Horrible About the Kind of Men Seeking Trad Wives To Get Swole, Teens Are Pumping Themselves Full of Drugs Meant for Fattening Cows for the Slaughterhouse Foolish Pollsters Are Now Just Asking AI What Voters Would Say in Response to Questions and Publishing It at Face Value OpenAI Says It’s Already Made $100 Million by Stuffing ChatGPT With Ads DOGE Made Drastic Cuts to a Global Vaccine Assistance Program. Now There’s a Deadly Measles Outbreak in Bangladesh Man Punished for Breaking Into Moo Deng’s Zoo Enclosure AI Is Causing Healthcare Costs to Surge There’s a Mass Rebellion Against AI in the Workplace People Who Lose Their Job to AI Are in for a World of Pain, Goldman Sachs Report Finds OpenAI Says Not to Worry About UBI, Because It Has Another Idea Police Officer Helplessly Waves Arms at Waymo That Careened Wrong Way Through Whataburger Drive-Thru Someone Just Threw a Molotov Cocktail At Sam Altman’s House New York Times Makes Substantial Changes to Article That Glazed a Sleazy AI Startup: “Our Piece Should Have Included That Information” Space Scientists Wince as Astronauts’ Lives Depend on Artemis 2’s Controversial Heat Shield During Plunge Back to Earth The Moon Astronauts Have Been Working Out With a NASA Rowing Machine in Space First AI Model From Zuckerberg’s Wildly Expensive Superintelligence Lab Flops Compared to Virtually All Rivals Economists Starting to Admit They May Have Been Wrong About AI Never Replacing Human Jobs AI-Powered Drug Marketer Medvi Responds After Allegations About Fake Doctors and Patients As Astronauts Visit the Moon, NASA Insider Says Agency Is in Shambles Behind the Scenes Man Lights 1.2 Million Square Foot Warehouse on Fire for Not Paying Him Enough NASA Scientists Screamed With Delight When They Saw Something Smashing Into the Moon Google Says Showing Polymarket Bets on Google News Was a Mistake Las Vegas Sphere Turns Into Huge Moon to Celebrate NASA Mission The New York Times Says It’s Identified the Creator of Bitcoin We Talked to a Writer Accused of Publishing An AI-Generated Essay in The New York Times Naked Man Bursts Into Tesla Service Center With a Shotgun Student Dies When Hospital Has No ICU Doctors, Calls One on Videochat Who Pronounces Him Dead Remotely, Lawsuit Claims
Analysis Finds That Google’s AI Overviews Are Providing M...
2026-04-08 · via Futurism

A man in a dark suit sits on a tall black stool facing away, wearing a tall white dunce cap. The background is bright yellow with a large red circle behind him.

Illustration by Tag Hartman-Simkins / Futurism. Source: DNY59 / Getty Images

Sign up to see the future, today

Can’t-miss innovations from the bleeding edge of science and tech

Google’s AI Overviews are peddling misinformation on a scale that may be virtually unprecedented in human history.

A recent analysis conducted by the AI startup Oumi at the behest of The New York Times found that the AI-generated summaries, which appear above Google search results, are accurate around 91 percent of the time. 

In a sense, that may sound like an impressive figure. But here’s an even more impressive one: five trillion. That’s roughly the number of search queries that Google processes every year, translating to tens of millions of wrong answers that the AI Overviews are providing every hour — and hundreds of thousands every minute, the analysis calculated.

In other words, Google has created a misinformation crisis. Studies have shown that people tend to trust what an AI tells them without question, with one report finding that only 8 percent of users actually double checked an AI’s answer. Another experiment found that users still listened to AI when it gave them the wrong answer nearly 80 percent of the time — a grim trend the researchers dubbed “cognitive surrender.”

Large language models adopt an authoritative tone and can confidently present fabricated information as fact when it can’t immediately glean a straight answer. Add the convenience that Google’s AI Overviews offer, and it’s easy to imagine untold numbers of users taking its summaries at their word.

Oumi conducted the analysis using a test called SimpleQA, a widely used benchmark for AI accuracy in the industry which was designed by OpenAI. The first round of tests, conducted in October, used a version of the AI Overviews powered by Google’s Gemini 2 model. A follow-up conducted in February tested the feature after it was switched to Gemini 3, its much-hyped upgrade.

Each round of tests involved 4,326 Google searches. Gemini 3 came out the more accurate model, giving a factually sound response 91 percent of the time. Gemini 2 performed significantly worse, at just 85 percent accurate.

On the one hand, it shows that the models are improving. On the other, it shows that Google was willing to foist a model on its userbase that was even more prone to hallucinating, in an ongoing experiment that’s still misinforming hundreds of millions of people.

Google called the analysis flawed. “This study has serious holes,” Ned Adriance, a Google spokesman, told the NYT in a statement. “It doesn’t reflect what people are actually searching on Google.”

Yet Google’s own tests paint a no less damning picture, the reporting notes. In an internal analysis of Gemini 3, it found that the AI model produced incorrect information 28 percent of the time. Google claims, however, that AI Overviews are more accurate because they draw on Google search results before answering.

The improvement between Gemini 2 and Gemini 3 may be papering over a more serious flaw. In the Oumi analysis, Gemini 2 provided answers that were “ungrounded” 37 percent of the time, meaning the AI Overviews cited websites that didn’t support the information they provided. But with Gemini 3, this jumped to 56 percent. On top of suggesting that the AI is pulling facts out of thin air, ungrounded responses make it difficult for users to verify the AI’s claims.

More on Google: Google News Now Prominently Featuring Polymarket Bets