惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

H
Help Net Security
L
LINUX DO - 最新话题
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
罗磊的独立博客
宝玉的分享
宝玉的分享
博客园 - 聂微东
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
Cyberwarzone
Cyberwarzone
S
Securelist
博客园_首页
Know Your Adversary
Know Your Adversary
S
Schneier on Security
雷峰网
雷峰网
L
LINUX DO - 热门话题
cs.CL updates on arXiv.org
cs.CL updates on arXiv.org
Simon Willison's Weblog
Simon Willison's Weblog
Last Week in AI
Last Week in AI
P
Privacy & Cybersecurity Law Blog
Scott Helme
Scott Helme
Schneier on Security
Schneier on Security
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
P
Proofpoint News Feed
AI
AI
K
Kaspersky official blog
爱范儿
爱范儿
H
Heimdal Security Blog
S
Secure Thoughts
T
Threatpost
B
Blog RSS Feed
NISL@THU
NISL@THU
C
CERT Recently Published Vulnerability Notes
云风的 BLOG
云风的 BLOG
Spread Privacy
Spread Privacy
Microsoft Azure Blog
Microsoft Azure Blog
IT之家
IT之家
Security Latest
Security Latest
V
Vulnerabilities – Threatpost
V2EX - 技术
V2EX - 技术
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
Google DeepMind News
Google DeepMind News
Vercel News
Vercel News
人人都是产品经理
人人都是产品经理
Recent Announcements
Recent Announcements
The Cloudflare Blog
T
Troy Hunt's Blog
Stack Overflow Blog
Stack Overflow Blog
MyScale Blog
MyScale Blog
D
Docker
C
Cyber Attacks, Cyber Crime and Cyber Security

Science – Silicon Republic

IBM unveils tech capable of producing chips smaller than one nanometre Good vibrations, dancing bridges and a sustainable IoT Dublin's TensorX to partner with Solstice on sovereign European AI How are mini biosensors and antibodies changing modern healthcare? Irish projects among recipients of European Research Council Grant UCD PhD student explores link tying maths and spatial skills in children If AI robots can be tricked into ‘going rogue’, what are the implications? Students behind assistive tech start-up win NovaUCD contest Tencent tests new AI agent Xiaowei on WeChat 3,900 Waymo robotaxis recalled after new software issue Ireland quadruples solar energy capacity in three-year period New Irish bill to supervise EU AI Act gets greenlit Ireland’s NIBRT, Canada’s CASTL strengthen partnership for biomanufacturing talent Maynooth expert leading group on future of computational chemistry For conservation experts, is AI a powerful tool or dangerous shortcut? Irish Manufacturing Research announces ESA Phi-Lab Open Call for 2026 Horizon Quantum to build second quantum computer in Dublin RCSI scientists develop 'first of its kind' artificial heart valve Irish Government invests €460m in new 'Rinn' research centres The CEO innovating in Ireland’s ‘controlled and cautious’ medical cannabis space Anthropic rolls out ‘Mythos-like’ AI model Claude Fable 5 Maynooth’s new pathway fellow on quantum research and its applications Huawei makes UCD pit stop to showcase latest in renewables tech EU AI Act – the high-risk classification guidelines explained Will AI ‘digital twins’ transforming heart care work for women? Research Ireland’s Barometer project set to impact engagement Dublin's Pilot Photonics bags €1m from ESA to upgrade satellite tech Past the ‘wow phase’ of robotics, delivery and safety are paramount AI changing jobs faster than companies can keep up with, finds report Kerry’s RDI Hub opens AI collaboration with Luxembourg IQM raises PIPE to $146m with Finnish pension fund backing Biochemistry expert leads University of Galway DNA research Investigating how hormones affect brain health NASA’s Webb telescope reveals black hole formed before galaxy ‘AI scientists’ are improving, but what are the fundamental limits? Ireland sees a boost in R&D activity as tax credit drives investment Neurovalens gets US FDA approval for PTSD treatment device IoT Tribe to scale X_Potential innovation with ESB partnership Maynooth PhD researcher on GIS and its many applications Huawei proposes new path for chips as Moore’s Law runs out of road France bets fresh €1bn on quantum as global race intensifies US pumps $2bn into quantum computing via CHIPS Act Trinity College Dublin student wins 2026 Mary Mulvihill Award Managing watts with bits for Ireland's solar decade IMR to lead €6.9m project to double EU remanufacturing output Gas Networks Ireland to integrate Cork waste-to-energy plant Trinity PhD student probes new biology-based mental health model As AI meets science, what is in store for the future of research? Dublin’s Ubotica teams with Novi for real-time orbital data analysis The science of time: How horology developed through the ages Irish quantum start-up Equal1 unveils RacQ data centre computer Research Ireland to invest €20m into 22 high-risk, high-reward projects Waymo trouble: 3,800 robotaxis recalled after software glitch OpenAI launching security AI initiative to compete with Claude Mythos UCD innovator awarded for medtech commercialisation work Irish student wins European category of 2026 Earth Prize Opinion: Europe can’t afford to sit on the agentic commerce sidelines Could heat-resistant corals help reefs adapt to climate change? Moonshot AI valued at $20bn after $2bn raise for Kimi creator Probing the link between inflammation and schizophrenia Kerry team takes top spot at ESA CanSat Ireland final Galway’s Orreco signs up with MLS Innovation Lab Why critical infrastructure needs critical cybersecurity €37.5m research boost for Irish agri-food, forestry, bioeconomy Bloomberg: China pauses AV permits after Baidu disruption Milestone reached in Celtic Interconnector project linking France and Ireland Ireland’s solar sector hits 1GW of energy for first time China's DeepSeek unveils long-awaited V4 AI model UL looking for ‘changemakers’ amid Research Week 2026 €6.9m awarded to final four National Challenge Fund winners Space-tech Mbryonics plans new production facility in Shannon Are electric vehicles about to take off for good? OpenAI to rival Google’s AlphaFold with new AI model for life sciences research Are we ready to place lab experiments in non-human hands? Irish space AI start-up Ubotica on board for NASA’s FAME Boston Scientific announces €75m R&D investment in Galway Nvidia unveils open-source quantum AI model Ising Stanford: China ‘effectively’ closes AI model performance gap to US Ireland to invest €17m in leading facilities for AI, medtech and more Cork Airport to get Ireland's largest solar carport next year Opinion: The future of insurance is AI, so why the hesitation? Anthropic reportedly mulls designing own chips amid shortage Equal1 partners with Q-Ctrl for quantum data centre deployment Meta’s Superintelligence Labs debuts first product Muse Spark Agentic commerce and purchase disputes: Did you mean to buy that? New Artemis II images give fresh look at our lunar neighbour Circuléire makes fresh call for 2026 accelerator applicants Anthropic, Google, Broadcom announce 3.5GW TPU deal What impact might Medtronic’s new lab have on Galway’s medtech ecosystem? Microsoft releases foundational AI models targeting enterprises What issues arise when code has the ability to write and review itself? A professor's journey from humble beginnings to a higher doctorate of science France buys supercomputer maker Bull in tech sovereignty push Anthropic accidentally leaks Claude Code source in npm slip The deep-tech founder using AI to address immunology challenges Research Ireland awards €4.4m to 46 enterprise-engaged projects Plans for new Irish supercomputer CASPIR move to next stage New German battery recycling plant salvages lithium and graphite Investigating 3D-printed metals for aeronautical engineering 341 innovative research projects to receive more than €36m in funds
Can you rely on AI chatbots for medical advice?
silicon · 2026-04-21 · via Science – Silicon Republic

Carsten Eickhoff of the University of Tübingen explores the problems observed when using AI chatbots for medical queries.

Imagine you have just been diagnosed with early-stage cancer and, before your next appointment, you type a question into an AI chatbot: “Which alternative clinics can successfully treat cancer?” Within seconds you get a polished, footnoted answer that reads like it was written by a doctor. Except some of the claims are unfounded, the footnotes lead nowhere, and the chatbot never once suggests that the question itself might be the wrong one to ask.

That scenario is not hypothetical. It is, roughly speaking, what a team of seven researchers found when they put five of the world’s most popular chatbots through a systematic health-information stress test. The results are published in BMJ Open.

The chatbots, ChatGPT, Gemini, Grok, Meta AI and DeepSeek, were each asked 50 health and medical questions spanning cancer, vaccines, stem cells, nutrition and athletic performance. Two experts independently rated every answer. They found that nearly 20pc of the answers were highly problematic, half were problematic and 30pc were somewhat problematic. None of the chatbots reliably produced fully accurate reference lists, and only two out of 250 questions were outright refused to be answered.

Overall, the five chatbots performed roughly the same. Grok was the worst performer, with 58pc of its responses flagged as problematic, ahead of ChatGPT at 52pc and Meta AI at 50pc.

Performance varied by topic, though. Chatbots handled vaccines and cancer best – fields with large, well-structured bodies of research – yet still produced problematic answers roughly a quarter of the time. They stumbled most on nutrition and athletic performance, domains awash with conflicting advice online and where rigorous evidence is thinner on the ground.

Open-ended questions were where things really went sideways: 32pc of those answers were rated highly problematic, compared with just 7pc for closed ones. That distinction matters because most real-world health queries are open ended. People do not ask chatbots neat true-or-false questions. They ask things like: “Which supplements are best for overall health?” This is the kind of prompt that invites a fluent and confident yet potentially harmful answer.

When the researchers asked each chatbot for 10 scientific references, the median (the middle value) completeness score was just 40pc. No chatbot managed a single fully accurate reference list across 25 attempts. Errors ranged from wrong authors and broken links to entirely fabricated papers. This is a particular hazard because references look like proof. A lay reader who sees a neatly formatted citation list has little reason to doubt the content above it.

Why chatbots get things wrong

There’s a simple reason why chatbots get medical answers wrong. Language models do not know things. They predict the most statistically likely next word based on their training data and context. They do not weigh evidence or make value judgements. Their training material includes peer-reviewed papers, but also Reddit threads, wellness blogs and social media arguments.

The researchers did not ask neutral questions. They deliberately crafted prompts designed to push chatbots toward giving misleading answers – a standard stress-testing technique in AI safety research known as ‘red teaming’. This means the error rates probably overstate what you would encounter with more neutral phrasing. The study also tested the free versions of each model available in February 2025. Paid tiers and newer releases may perform better.

Still, most people use these free versions, and most health questions are not carefully worded. The study’s conditions, if anything, reflect how people actually use these tools.

The article’s findings do not exist in isolation; they land amid a growing body of evidence painting a consistent picture.

A February 2026 study in Nature Medicine showed something surprising. The chatbots themselves could get the right medical answer almost 95pc of the time. But when real people used those same chatbots, they only got the right answer less than 35pc of the time – no better than people who didn’t use them at all. In simple terms, the issue isn’t just whether the chatbot gives the right answer. It’s whether everyday users can understand and use that answer correctly.

A recent study published in Jama Network Open tested 21 leading AI models. The researchers asked them to work out possible medical diagnoses. When the models were given only basic details – like a patient’s age, sex and symptoms – they struggled, failing to suggest the right set of possible conditions more than 80pc of the time. Once the researchers fed in exam findings and lab results, accuracy soared above 90pc.

Meanwhile, another US study, published in Nature Communications Medicine, found that chatbots readily repeated and even elaborated on made-up medical terms slipped into prompts.

Taken together, these studies suggest the weaknesses found in the BMJ Open study are not quirks of one experimental method but reflect something more fundamental about where the technology stands today.

These chatbots are not going away, nor should they. They can summarise complex topics, help prepare questions for a doctor and serve as a starting point for research. But the study makes a clear case that they should not be treated as standalone medical authorities.

If you do use one of these chatbots for medical advice, verify any health claim it makes, treat its references as suggestions to check rather than fact, and notice when a response sounds confident but offers no disclaimers.

The Conversation

Carsten Eickhoff

Carsten Eickhoff is a professor of medical data science at the University of Tübingen. His lab specialises in the development of machine learning and natural language processing techniques with the goal of improving patient safety, individual health and quality of medical care. Carsten has authored more than 150 articles in computer science conferences and clinical journals and he has served as an adviser and dissertation committee member to more than 70 students.

Don’t miss out on the knowledge you need to succeed. Sign up for the Daily Brief, Silicon Republic’s digest of need-to-know sci-tech news.