惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

GbyAI
GbyAI
The GitHub Blog
The GitHub Blog
小众软件
小众软件
美团技术团队
博客园 - 司徒正美
G
Google Developers Blog
Blog — PlanetScale
Blog — PlanetScale
Hugging Face - Blog
Hugging Face - Blog
博客园_首页
大猫的无限游戏
大猫的无限游戏
罗磊的独立博客
Recent Announcements
Recent Announcements
酷 壳 – CoolShell
酷 壳 – CoolShell
D
Docker
J
Java Code Geeks
Last Week in AI
Last Week in AI
V
Visual Studio Blog
Microsoft Azure Blog
Microsoft Azure Blog
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
P
Proofpoint News Feed
V
V2EX
C
Check Point Blog
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
MyScale Blog
MyScale Blog

Fortune | FORTUNE

One man can kill Bill Ackman’s $64 billion bid for Universal Music Group—and no one knows what he’ll do | Fortune Poppi’s cofounder pitched her startup on Shark Tank while 9 months pregnant and landed a $400,000 deal—now it's worth $2 billion | Fortune Teen boys are choosing AI girlfriends over real ones for 'maximum control, zero rejection'—experts say it could make them unemployable | Fortune A United American merger is by no means impossible given the president 'loves big deals' | Fortune Reed Hastings’s planned exit from $455 billion Netflix ‘had nothing to do with’ the failed deal for Warner Bros., says Ted Sarandos | Fortune Meet Joe McCann: The high-flying crypto trader held in Tanzania after sudden death of his influencer fiancée Ashly Robinson | Fortune Gen Z is carving a different path in the housing market by doing it alone | Fortune U.S. Catholic leaders criticize Trump for ‘disparaging words’ about the pope as Vatican clash risks alienating Catholic voters | Fortune China has ‘nearly erased’ America’s lead in AI—and the flow of tech experts moving to the U.S. is slowing to a trickle, Stanford report says | Fortune Self-made millionaire behind $5 billion Skims Emma Grede says it all began with a cold call to Kris Jenner: Emma Grede—the self-made millionaire behind the $5 billion Skims empire—says it all began with an audacious cold call to Kris Jenner: ‘The difference between me and someone else is, I made it happen’ | Fortune Americans have never been this gloomy about the economy. Wall Street has never cashed in harder | Fortune ‘The college grading system [is] almost meaningless’: People see the Ivy League as an easy A and with flawed admissions standards | Fortune The CEO of $8.5 billion Japanese car giant Nissan plays the drums in a band and hits the tennis courts to destress from the top job | Fortune New York governor's take on a millionaires tax: fancy pied-à-terre second apartments worth over $5 million | Fortune Pope Leo XIV: A ‘handful of tyrants’ are ravaging earth with war and exploitation | Fortune Trump has no plan to cut the $39 trillion national debt, but he does want to cut childcare. His budget director is scrambling to clarify | Fortune China's economy grows 5% in first quarter, surprising economists to the upside | Fortune Everyone was wondering what Trump wanted more: Warsh smoothly seated at the Fed, or for Powell to pay. We have our answer | Fortune Palantir exec: the biggest mistake retailers are making with AI? Trying to do it all with one agent | Fortune American YouTuber who calls himself a 'troll' sentenced to 6 months in Korean prison for literally dancing on wartime graves | Fortune BBC plans to cut up to 2,000 jobs to save 10% of annual budget | Fortune Canva debuts a new suite of agentic tools, as the design app quietly becomes one of the world’s most used AI services | Fortune Moody's CEO: AI has a trust problem – better models won’t fix it | Fortune Top New York surgeon: Americans have better data for choosing restaurants than surgeons. That has to change | Fortune The Iran war’s fertilizer shock is hammering American farmers, and 70% can’t afford what they need for this year’s growing season | Fortune Education experts to Mamdani: Why are you foisting AI on our kids? | Fortune This CEO pirated video games as a teen and became a hacker for the Air Force. Now he’s built a $3 billion cyber firm | Fortune Teacher, blame thyself: Yale report savages Ivy League schools for destroying American trust in higher education | Fortune Fed chair nominee Kevin Warsh is worth more than $100 million and has stakes in SpaceX and Polymarket | Fortune From wool sneakers to GPUs: Allbirds’ desperate AI pivot and 600% stock surge, explained | Fortune
AI hallucinations are slipping past experts into papers a...
Tristan Bove · 2026-05-24 · via Fortune | FORTUNE

It was a process that had become routine for Maxim Topaz. 

The associate professor at Columbia University’s School of Nursing had grown accustomed to having artificial intelligence tools help polish scientific papers for grammar, formatting, and other details. But a few weeks after submitting his latest research, the academic journal he was due to publish in came back with questions about a reference. The AI tool Topaz had used had silently inserted a fabricated source into his work.

“I felt deeply embarrassed,” Topaz, who leads a team at Columbia developing AI applications in healthcare, told Fortune

“I’m an AI researcher. I know about hallucinations,” he said. “If this is happening to me, an AI expert, what happens to other people?”

That near-miss sent Topaz on an investigation to find out how often experts were getting subtly fooled by AI. The answer, it turns out, is a lot. 

In a study published earlier this month in The Lancet, Topaz and his colleagues audited nearly 2.5 million biomedical papers and 97 million citations indexed on PubMed Central, the central repository used by clinicians and researchers worldwide. They found more than 4,000 fabricated references buried across nearly 3,000 papers. Not all the references were AI-generated, though Topaz said the steady rise in fake sourcing went “vertical” in 2024, shortly after AI tools in research entered more widespread use.

“It’s very reasonable that AI is highly associated with them now,” he said.

Over the past three years, the rate of fabricated references in biomedical literature has grown more than 12-fold. In 2023, one in 2,828 papers contained at least one fake reference, a rate that had risen to one in 458 by last year. Over the first seven weeks of 2026, the researchers found, one in 277 papers had at least one non-existent reference. 

“I’m thinking this is just the tip of the iceberg,” Topaz said.

Hallucinations happen when an AI model prioritizes word patterns over accuracy. They are often harmless, but the stakes are different when AI errors begin infiltrating academic literature, as hallucinations risk undermining the scientific process. 

Medicine is a field that builds on itself. Clinical trials cite earlier studies; systematic reviews then aggregate those trials, and medical guidelines finally cite those reviews. Doctors and nurses rely on those guidelines when they decide how to treat patients. A fabricated study planted at the start of that process doesn’t stay there.

“This is the evidence chain, that’s how we care for and treat people. If you put the fictional study at the bottom of the stack, the whole structure inherits it,” Topaz said. 

“We’ve already seen paper mill articles included in systematic reviews informing clinical guidelines,” he added. “When a guideline paper cites a paper with a partially fictional references list, the evidence-based chain for treatment decisions is compromised.”

AI mistakes come for everyone

That AI is vulnerable to hallucinations has been known since ChatGPT first entered the scene four years ago, when students began to bravely submit specious AI-generated papers under their own name. But with a litany of tools, agents, and extensions now ubiquitous in nearly every profession, even experts in their field are getting tripped up by AI.

Take the case of Steven Rosenbaum. The author and filmmaker was in the headlines for all the wrong reasons this week after the New York Times identified a slew of inaccurate quotes throughout his new book, titled The Future of Truth: How AI Reshapes Reality

The book carried blurbs from prominent journalists, including Nicholas Thompson, The Atlantic’s chief executive, and a foreword by Maria Ressa, the Nobel Peace Prize–winning reporter from the Philippines. It arrived, according to the Times, “to great fanfare.”

Rosenbaum’s book contained more than a half-dozen misattributed or entirely invented quotes, apparently generated by AI tools he had disclosed using in his acknowledgments. In a statement to the Times, Rosenbaum recognized the errors, calling the episode “a warning about the risks of AI-assisted research and verification.”

Instances like these might be inevitable given how widely AI is being used in expert-level knowledge work. Several journalism outlets, Fortune included, are now piloting the use of AI tools in reporting. Surveys suggest more than half of legal professionals are using AI tools to draft briefs and memos. A recent report by the American Medical Association found over 80% of physicians now use AI professionally to summarize research and prepare clinical documentation, a share that has more than doubled since 2023. Even Nobel laureates, such as Literature Prize winner Olga Tokarczuk, admit to using AI in their work.

As for research, one study last year by an American medical journal identified 36% of its papers contained at least some AI-generated text, although only 9% of researchers disclosed this when prompted prior to submitting their manuscripts. Another recent study found more than half of researchers are likely to be using AI tools while peer-reviewing other people’s work.

But as it turns out, experts in their field are no less likely to get duped. Topaz’s study of hallucinations in biomedical research joins a growing pile of anecdotes and datasets documenting embarrassing errors, including legal analyst Damien Charlotin’s catalog of 1,459 legal decisions citing AI-generated inaccurate content. Before he started the project a year ago, AI hallucinations in legal cases appeared two or three times a month. Now, there’s around five a day.

When experts get it wrong

Fake AI-generated research papers are already a problem in academia, increasingly difficult to parse through and threatening to overwhelm the peer-review system. But hallucinated references in real studies produced by humans could be just as widespread, and potentially even harder to track down.

The vast majority of papers tracked by Topaz contained only one or two fabricated citations, out of the several dozen references academic studies usually need to publish, suggesting most cases of AI hallucinations in research are unintentional. 

But the publishing industry might not be prepared to handle the surging number of fake references, Topaz said. Verification methods differ between journals, and while some use software to check references and scan for AI-generated content, enforcement varies wildly. There is also no easy mechanism to retroactively screen the evidence chain to find original fake studies or references. So far, few journals have been able to identify hallucinations, as Topaz’s analysis found 98.4% of studies with fake references had not been retracted by publishers at the time of his audit.

It’s part of what people in the field have referred to as science’s “reproducibility crisis,” compounded in the age of AI by a rising flood of useless or unreliable AI-generated content that now permeates academic literature. But it’s a similar story in other fields that rely on output that can be reproduced. Stories in newspapers drive conversations and form the bedrock of future investigations. Legal decisions are eventually cited by lawyers and scholars in other cases. 

Topaz said AI itself is not necessarily the villain, and he gladly uses it in his own work. “The problem is unverified AI output entering the permanent record,” he said. “The fix is not to stop using the tools, it’s to build verification into the workflow.”

“The longer we wait to put verifications in place, the harder it becomes to clean up,” he added.

AI hallucinations don’t care how well-versed in a subject users are. The mistakes are designed to look real, and they’re getting better at hiding. The more consequential the field—be it medicine, law, or journalism—the more dangerous errors become when they aren’t caught.