惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

S
SegmentFault 最新的问题
博客园 - 三生石上(FineUI控件)
WordPress大学
WordPress大学
博客园 - 【当耐特】
月光博客
月光博客
Vercel News
Vercel News
D
Docker
I
InfoQ
Apple Machine Learning Research
Apple Machine Learning Research
博客园 - 叶小钗
MongoDB | Blog
MongoDB | Blog
GbyAI
GbyAI
有赞技术团队
有赞技术团队
雷峰网
雷峰网
博客园 - 聂微东
小众软件
小众软件
Y
Y Combinator Blog
腾讯CDC
L
LangChain Blog
The GitHub Blog
The GitHub Blog
宝玉的分享
宝玉的分享
Stack Overflow Blog
Stack Overflow Blog
大猫的无限游戏
大猫的无限游戏
T
The Blog of Author Tim Ferriss

Latest from Live Science

Loneliness may contribute to memory issues, but not dementia — they are 'not the same thing' Naked mole rats wage bloody wars of succession to choose a new queen — but one colony did something scientists… Can the US be trusted with the moon? A law scholar raises concerns after Artemis II Lyrid meteor shower 2026: See spring's first rain of 'shooting stars' peak in moonless skies $3 million prize goes to duo whose research led to first sickle cell CRISPR therapy 700-year-old mummy from Bolivia contains earliest confirmed evidence of strep throat bacteria in the Americas New pain-relief opioid could be much less addictive than morphine, rodent study finds Experimental drug doubles one-year survival in pancreatic cancer Science news this week: Physicists witness faster-than-light darkness pinpricks, humans are still evolving, and some… Archaeologists discover perfectly circular ancient Egyptian temple that may have been used for sacred water rituals Some polar bears are adapting to their melting habitat. Will it be enough to save the iconic species? 2 supermassive black holes may collide 100 years from now ‪—‬ and Earth would feel it Anglo-Saxon burial holds an older sister cradling her little brother after they both died 1,400 years ago, possibly of… Colorado River may have pooled and spilled over to form the Grand Canyon, solving a long-standing mystery ‪—‬… 'We all screamed when it happened': Bright-green fireball meteor caught exploding over famous Viking raid site… Northern lights may be visible from several US states Friday and Saturday as giant hole opens up in sun Hackers used AI to steal hundreds of millions of Mexican government and private citizen records in one of the largest… The first black hole ever discovered is spewing 'dancing jets' at half the speed of light Stephen Hawking's black hole information paradox could be solved — if the universe has 7 dimensions 'Something's missing': Most thorough-ever study of the cosmos proves we still can't explain how the… 'Human evolution didn't slow down; we were just missing the signal': Large DNA study reveals natural selection led to more redheads and less male-pattern baldness Artemis II quiz: Is your knowledge of NASA New study confirms lobsters feel pain, driving scientists to call for a ban on boiling them alive This humanoid robot does all your housework for you ‪—‬ and its makers say it Ancient process that created rare earth elements discovered — and it could help us locate desperately needed deposits Strange mammal ancestor laid huge, leathery eggs —‬ and it was key to surviving the world 73 moon landings? NASA Diagnostic dilemma: A woman heard voices telling her she had a brain tumor ‪—‬ and scans confirmed she did Triassic croc relative from Ghost Ranch, New Mexico finally identified after nearly 80 years in museum basement There were
AI 'mirages' mean tools used to analyze medical scans cou...
2026-04-07 · via Latest from Live Science
A man with clear glasses wearing a white lab coat and stethoscope looks at a holographic blue and orange image of a leg and leg bone.
AI models are being trained to interpret medical scans, but researchers warn that a flaw in these systems could undermine their accuracy. (Image credit: Westend61 via Getty Images)

Researchers have been training artificial intelligence (AI) systems to interpret results of visual tests like mammograms, MRIs and tissue biopsies — and as AI becomes increasingly capable, some analysts have suggested that these models will replace humans in the field of medical diagnostics.

But now, a new study casts doubt on the capability of current AI models to deliver reliable results, highlighting a crucial flaw that could hinder their use in medicine.

They called this phenomenon a "mirage," and it is the first time this effect has been shown across multiple AI models, which were used to interpret images across multiple disciplines.

"What we show is that even if your AI is describing a very, very specific thing that you would say, 'Oh, there's no way you could make that up,' yeah, they could make that up," said study first author Mohammad Asadi, a data scientist at Stanford University. "They could make very rare, very specific things up."

When AI sees what isn't there

AI "hallucinations" are well documented and involve models filling in made-up details, such as false citations for a real essay. They often result from AI making inaccurate or illogical predictions based on training data it was provided. The scientists instead called the phenomenon in the new study "mirages" because the AI created descriptions of original images on their own and then based their answers on those nonexistent images.

In the study, the researchers gave 12 models a text input prompt, such as "Identify the type of tissue present in this histology slide." Then, they either provided the image of the slide or they did not. When a model was not provided with an image, sometimes it would alert the human user that no image was provided. However, most of the time, the model would instead describe an image that did not exist and provide an answer to the original prompt.

Get the world’s most fascinating discoveries delivered straight to your inbox.

The researchers observed this "mirage mode" across 20 disciplines, testing models' interpretations of a variety of images, from satellites to crowds to birds. The mirage effect was seen across all the disciplines and all the AI models, to varying levels. But it was particularly pronounced in medical diagnostics.

When given text prompts about brain MRIs, chest X-rays, electrocardiograms or pathology slides, but no actual images, the AI models' answers also tended to be biased toward diagnoses that required immediate clinical follow-up. So, if used for clinical decision-making, the AI might prompt more aggressive medical care than is required, the team concluded.

Why AI invents images

So how does an AI model describe images that don’t exist?

The models, which have been trained on massive amounts of textual and visual data, aim to find the answer to a question in the fewest steps possible. And they will take whatever shortcuts they can to deliver an answer, studies have shown. Thus, models can end up relying solely on this trained logic rather than on provided images.

Digital mind, brain, artificial intelligence concept.

AI models could be powerful tools to improve medical diagnostics. But their inner workings aren't yet fully understood, and that can lead to assumptions about how well they analyze images. (Image credit: BlackJack3D via Getty Images)

Interestingly, when in mirage mode, AI models also perform well against benchmark tests typically used to assess their accuracy, the researchers found. These standardized tests challenge a model to complete a task — like answering multiple-choice questions — and compare its performance against an answer key of expected outputs.

Researchers can tweak the benchmark tests to assess an AI's visual understanding of images, but this approach doesn't account for questions answered based on mirages. Additionally, AI models are often trained on the same data that's used as a reference to write the benchmark tests. So it's possible for a model to answer questions based on that reference data, rather than by actually interpreting images.

According to Asadi, this is a problem because there is no way to tell whether an AI model has actually analyzed an image or is just making things up. If you are uploading a bunch of images but a few are corrupt or otherwise missing from the dataset, the model may not tell you. And it could still provide very coherent, comprehensive and convincing answers based on mirage images.

"[AI models] are very good at interpreting images," Asadi said. "But on the other hand, they're also very, very good at convincing us of things … and talking to us in an authoritative way."

That authority is apparent in the fact that many consumers query AI chatbots for health guidance, with about one-third of U.S. adults reporting that they do so. This conversational authority increases the risk that fabricated or overconfident outputs are trusted by both the general public and medical professionals, the study authors say.

RELATED STORIES

"We urgently need a new generation of evaluation frameworks that strictly measure true cross-modal integration — ensuring the AI is truly 'seeing' the pathology rather than just 'reading' the clinical context," Hongye Zeng, a biomedical AI researcher in the department of radiology at UCLA who was not involved in the study, told Live Science in an email.

This study shows that, while AI has become an increasingly useful tool in medical diagnostics, there are still aspects of its inner workings that we don't understand. Adasi thinks AI models can spot things that may be missed by medical professionals, but he also believes there should be a limit to how much we trust them.

AI companies have attempted to raise guardrails to prevent their models from hallucinating or spreading misinformation — but even these safeguards won't completely prevent the mirage effect, Asadi cautioned.

Jennifer Zieba earned her PhD in human genetics at the University of California, Los Angeles. She is currently a project scientist in the orthopedic surgery department at UCLA where she works on identifying mutations and possible treatments for rare genetic musculoskeletal disorders. Jen enjoys teaching and communicating complex scientific concepts to a wide audience and is a freelance writer for multiple online publications.

You must confirm your public display name before commenting

Please logout and then login again, you will then be prompted to enter your display name.