惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Threat Intelligence Blog | Flashpoint
Threat Intelligence Blog | Flashpoint
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
Recent Announcements
Recent Announcements
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
A
About on SuperTechFans
I
InfoQ
F
Full Disclosure
美团技术团队
Martin Fowler
Martin Fowler
量子位
V
V2EX
小众软件
小众软件
爱范儿
爱范儿
宝玉的分享
宝玉的分享
aimingoo的专栏
aimingoo的专栏
有赞技术团队
有赞技术团队
F
Fortinet All Blogs
M
MIT News - Artificial intelligence
T
Tailwind CSS Blog
博客园 - 三生石上(FineUI控件)
N
Netflix TechBlog - Medium
大猫的无限游戏
大猫的无限游戏
Vercel News
Vercel News
IT之家
IT之家
云风的 BLOG
云风的 BLOG
B
Blog RSS Feed
H
Hackread – Cybersecurity News, Data Breaches, AI and More
Hugging Face - Blog
Hugging Face - Blog
Microsoft Azure Blog
Microsoft Azure Blog
T
Threat Research - Cisco Blogs
T
Threatpost
P
Proofpoint News Feed
腾讯CDC
博客园 - 司徒正美
Jina AI
Jina AI
The Hacker News
The Hacker News
P
Privacy & Cybersecurity Law Blog
L
LINUX DO - 热门话题
S
Securelist
U
Unit 42
T
The Exploit Database - CXSecurity.com
博客园 - Franky
NISL@THU
NISL@THU
D
Docker
The GitHub Blog
The GitHub Blog
Latest news
Latest news
S
Schneier on Security
J
Java Code Geeks
Blog — PlanetScale
Blog — PlanetScale
P
Palo Alto Networks Blog

Latest from Live Science

Loneliness may contribute to memory issues, but not dementia — they are 'not the same thing' Naked mole rats wage bloody wars of succession to choose a new queen — but one colony did something scientists… Can the US be trusted with the moon? A law scholar raises concerns after Artemis II Lyrid meteor shower 2026: See spring's first rain of 'shooting stars' peak in moonless skies $3 million prize goes to duo whose research led to first sickle cell CRISPR therapy 700-year-old mummy from Bolivia contains earliest confirmed evidence of strep throat bacteria in the Americas New pain-relief opioid could be much less addictive than morphine, rodent study finds Experimental drug doubles one-year survival in pancreatic cancer Science news this week: Physicists witness faster-than-light darkness pinpricks, humans are still evolving, and some… Archaeologists discover perfectly circular ancient Egyptian temple that may have been used for sacred water rituals Some polar bears are adapting to their melting habitat. Will it be enough to save the iconic species? 2 supermassive black holes may collide 100 years from now ‪—‬ and Earth would feel it Anglo-Saxon burial holds an older sister cradling her little brother after they both died 1,400 years ago, possibly of… Colorado River may have pooled and spilled over to form the Grand Canyon, solving a long-standing mystery ‪—‬… 'We all screamed when it happened': Bright-green fireball meteor caught exploding over famous Viking raid site… Northern lights may be visible from several US states Friday and Saturday as giant hole opens up in sun Hackers used AI to steal hundreds of millions of Mexican government and private citizen records in one of the largest… The first black hole ever discovered is spewing 'dancing jets' at half the speed of light Stephen Hawking's black hole information paradox could be solved — if the universe has 7 dimensions 'Something's missing': Most thorough-ever study of the cosmos proves we still can't explain how the… 'Human evolution didn't slow down; we were just missing the signal': Large DNA study reveals natural selection led to more redheads and less male-pattern baldness Artemis II quiz: Is your knowledge of NASA New study confirms lobsters feel pain, driving scientists to call for a ban on boiling them alive This humanoid robot does all your housework for you ‪—‬ and its makers say it Ancient process that created rare earth elements discovered — and it could help us locate desperately needed deposits Strange mammal ancestor laid huge, leathery eggs —‬ and it was key to surviving the world 73 moon landings? NASA Diagnostic dilemma: A woman heard voices telling her she had a brain tumor ‪—‬ and scans confirmed she did Triassic croc relative from Ghost Ranch, New Mexico finally identified after nearly 80 years in museum basement There were Physicists witness pinpricks of darkness moving faster than the speed of light ‪—‬ without breaking the laws of relativity Mini lake meets snowy rim of Canada's oldest ice mass — Earth from space Stone Age tombs in Scotland reveal 'webs of descent' among male relatives 'Oslo patient' likely cured of HIV after getting stem cell transplant from his brother, who is genetically… Antiseptic-tolerant germs spread through the air in hospitals, early study hints Homo erectus' tools include stunning geodes and fossils, possibly as a way to connect with the cosmos, study finds 'Really, really weird': Physicists entangle two moving atoms for the first time, validating 'spooky'… www.livescience.com Sperm quality is at its peak in the summer, study finds Scientists are trying to build a vaccine that works against almost any respiratory pathogen  — here's… Idol of Pomos: A 5,000-year-old fertility figurine from Cyprus that wears a miniature version of herself on a necklace Human ancestors butchered and ate elephants 1.8 million years ago, helping to fuel their large brains Ancient Egyptian stone monument depicting a Roman emperor as a pharaoh discovered in Luxor 'Human minds shouldn't have to go through' this: Artemis II crew recalls unreal moment when Earth disappeared — Space photo of the week Does the moon look the same from everywhere on Earth? I found a new meteor shower — and it comes from an asteroid getting baked to bits by the sun AI for breakup texts? How 'sycophantic' chatbots are messing with our ability to handle difficult social… Science news this week: Artemis II splashes down, the world's fattest parrot bounces back, and the Shroud of Turin… 10 Artemis II photos that define humanity's return to the moon Do the microbes in your gut influence what foods you like? 'I'm at a loss for words': Artemis II mission comes home to joy and cheers after historic 10-day mission There are 'reasons to be confident' about faulty Artemis II heat shield ahead of 25,000 mph reentry, space… The moon is green and brown? Why scientists are already excited about Artemis II's historic lunar photos 'I've seen the movies. What a horrible way to die': What it's like to be sucked into a tornado and… 'More questions than answers': Experts baffled by Alaskan mammal-eating orcas spotted near Seattle Changing 'just one DNA letter' in female mice triggers growth of male genitalia Aoshima: Japan's tiny 'Cat Island' where felines hugely outnumber humans 'Welcome home, Integrity': Artemis II crew return to Earth after 'bullseye landing' caps historic… AI war games almost always escalate to nuclear strikes, simulation shows Ancient Korean society practiced human sacrifice and high inbreeding, researchers find There's an issue with the Artemis II heat shield, but NASA isn't worried. Here's why. Chimpanzees in Uganda are locked in a deadly 'civil war' after their group split apart — and scientists… James Webb telescope spots 'stingray' galaxy system that could solve the mystery of 'little red… 'RIP, Comet MAPS': Watch the superbright sungrazer become a 'headless wonder' after being ripped… Scientists create new type of encryption that protects video files against quantum computing attacks Western states face above-normal wildfire threats this summer. New maps reveal which areas are most at risk. Science history: Doctor hypothesizes that 'transmissible proteins' can cause disease, contradicting a 'central dogma' of molecular biology — April 9, 1982 Keratin may act as a 'brake' for skin inflammation, pointing to potential treatments 'No one knows what they are': Researchers discover new type of cell that's seen only during pregnancy 16th-century silver coin discovered near Strait of Magellan marks the spot of a doomed Spanish colony How to see Comet PanSTARRS as it brightens in the night sky this week Diagnostic dilemma: Woman's 'biologically implausible' infection led her to sneeze 'worms' out… 'In every continent where humans are present, water bankruptcy is manifesting itself': Exiled Iranian scientist Kaveh Madani on our desperate need to preserve our most precious resource California declared war on smog in the 1970s. The knock-on effects were huge. 'They are literally everywhere': The shocking story of how forever chemicals polluted the world DNA reveals ancestry of man buried in Stone Age monument in Spain, but his religion remains a mystery 'So much magic': Artemis II shares first images from the far side of the moon, including new… AI 'mirages' mean tools used to analyze medical scans could fabricate their findings World's fattest parrot — on the verge of extinction 30 years ago — has record-breaking breeding season It's one of the best toothbrushes we have tested (and it's not Oral-B) Physicists moved volatile antimatter by truck for the first time ever — paving the way for groundbreaking new… Deadly, vivid-green mass sprawls across South African reservoir — Earth from space The Artemis II astronauts have just flown farther from Earth than any humans in history Artemis II moon flyby begins: How to watch and what to know 'A cure on the horizon': Are we finally close to ending type 1 diabetes? 'They could spend 4 or 5 hours per day underwater': How humans adapted to the most challenging environments We went to Finland to hear about the new 'sand battery' that will turn stored renewable energy back into power… The hungriest black holes in the universe are running out of food, survey of 8,000 cosmic monsters reveals Beadnet dress: A 4,500-year-old ancient Egyptian funeral 'gown' that was in vogue during the Old Kingdom 'This generation's moment': How the Artemis missions will reframe humanity's relationship with the… Antarctica hides huge caches of gold, silver, copper and iron. As the ice melts, countries may race to harvest them. NASA telescope uncovers new mystery in supernova first spotted by Chinese astronomers 2,000 years ago —‬ Space… Diabetes rates are lower in high-altitude environments ‪‪—‬ and scientists may have discovered why Shroud of Turin, claimed to be Jesus' burial cloth, contaminated with carrot and red coral DNA What happened to the Minoan civilization? I've witnessed nearly 100 rocket launches. Artemis II was like nothing I've ever experienced. Science news this week: Artemis II lifts off, diabetes cured in mice, and smog in China shapes Arctic storms Fossil site in China reveals bevy of complex creatures lived prior to the Cambrian explosion, including a… Cheap, decades-old transplant drug delays full onset of type 1 diabetes Octopus quiz: Are you a sucker for cephalopod science?
AI-written code can beat humans at biomedical analysis, some studies find. What does that mean for the field?
2026-04-06 · via Latest from Live Science
A woman with dark straight hair pulled back wearing navy blue scrubs and a stethascope taps on a glass panel lit up with various technological images
Large language models can be a force multiplier for medical researchers but not without well-defined guardrails or humans in the loop. (Image credit: Krongkaew via Getty Images)

As the general public has embraced large language models (LLMs) such as ChatGPT, Claude and Gemini, scientists have been exploring how these artificial intelligence (AI) tools could enhance medical research.

Some argue that LLMs could dramatically boost researchers' efficiency in completing certain types of medical studies, and research published in February in the journal Cell Reports Medicine exemplifies that vision for the technology.

The study used massive datasets of patient biomedical information to predict the risk of preterm birth in a given pregnancy. These types of predictions have been a powerful AI use case for years, and were possible with more traditional types of machine learning than LLMs employ. But this study was notable in that LLMs enabled junior researchers — a graduate student and a high school student — to efficiently generate very accurate code.

That code predicted a baby's gestational age at birth and the likelihood of preterm birth. The AI's output matched and, in one case, even beat analyses from expert teams who had used human-generated code to crunch the same data.

"What I saw with junior scientists here and how effective they could be truly inspired and amazed me," said study co-author Marina Sirota, interim director of the Baker Computational Health Sciences Institute at the University of California, San Francisco.

One big promise of LLMs is to lower the barrier for researchers to produce code and conduct complex analyses — but it comes with risks. As AI quickly improves, researchers must grapple with myriad questions. What guardrails need to be established to ensure AI's accuracy? How do we measure its output? And how will the role of human researchers evolve as these systems gain prominence?

How AI prediction works

Sirota's team drew on data used in the Dialogue for Reverse Engineering Assessments and Methods (DREAM) Challenges, international competitions in which teams of scientists tackle complex biomedical problems using shared datasets.

Get the world’s most fascinating discoveries delivered straight to your inbox.

The open-source datasets included blood transcriptomics, which looks at RNA, a molecule that reflects which genes are active in the body. They included epigenetic information from placental cells, which described chemical tags that sit "on top of" DNA and control which genes can be switched on, and microbiome data describing the bacteria present in vaginal fluid samples.

These data points were flagged with the type of sample they came from — blood, placental tissue or vaginal fluid — and labeled with outcomes of interest, namely gestational age and preterm birth. Machine learning algorithms can then be trained to spot links between a sample's source and its label. For example, they may reveal that microbiome samples with certain mixes of bacteria often come from people who have given birth early.

Once trained on a subset of data, the algorithm can be tested on samples that lack labels, to see if it can predict the label that should be there. For instance, it should flag samples with bacterial mixes similar to those in the training data linked to a higher risk of preterm birth.

But we can speed that up as well — the cleaning part and normalization of data — with generative AI.

Marina Sirota, interim director of the Baker Computational Health Sciences Institute at the University of California, San Francisc

The final step is to evaluate the models' accuracy and compare them. "Accuracy" in the context of machine learning has a specific definition: the number of correct predictions divided by the total number of predictions.

Human- vs. AI-generated code

The DREAM Challenge was aimed at uncovering links between these medical metrics and the risk of preterm birth. Some risk factors, including having infections during pregnancy, are already well known. But the DREAM Challenge wanted to see what signals might be gleaned from clinical samples, like blood.

It's the kind of work that normally demands months of effort from trained bioinformaticians. But instead of writing the analysis code themselves, the junior researchers in the recent study gave each of eight LLMs a single prompt describing the data available and the labeling task at hand: predicting gestational age or preterm birth.

LLMs tested

  • ChatGPT o3-mini-high
  • ChatGPT 4o
  • DeepSeek R1
  • Gemini 2.0 FlashExpThink
  • Qwen 2.5 Coder
  • Llama 3.2
  • Phi-4
  • DeepSeek-R1-Distill-Qwen

With this simple prompting, four of the eight models — DeepSeekR1, Gemini, and ChatGPT's o3-mini-high and 4o — produced code that ran successfully. The best performer, OpenAI's o3-mini, was as accurate as the original human DREAM Challenge teams. For one task, which involved estimating gestational age from epigenetic data, it was more accurate than humans had been.

What's more, the junior researchers generated results in about three months and submitted a manuscript describing their results within six months, whereas the same process took the original DREAM Challenge teams years.

"We got lucky with the review process here, but six months to generate the results and write the paper is pretty incredible, especially for a junior scientist," Sirota told Live Science.

Preterm birth, before 37 complete weeks of pregnancy, affects roughly 11% of infants worldwide. Babies born too early are at higher risk than full-term babies for a host of health troubles, including but not limited to problems affecting their brains, eyes and digestive systems. Being able to predict which pregnant patients are more likely to give birth early could mean closer monitoring and treatments to protect the baby and make full-term birth more likely, experts say.

Beyond writing code

The data used in the Cell Reports Medicine paper started "in good shape," Sirota noted, in tables that AI could easily read. "But we can speed that up as well — the cleaning part and normalization of data — with generative AI," she said.

Sirota's team is now exploring other LLM applications, including a new tool called Chat PTB (short for "preterm birth") that they've developed. The Chat GPT-based tool is embedded in papers published by the March of Dimes research network, part of a nonprofit aimed at improving maternal and infant health. Instead of manually combing through this literature, researchers can now query Chat PTB and get synthesized answers with references — a task that used to take hours, compressed into seconds.

But tools like Chat PTB and the code-writing approach in Sirota's study represent only the first wave. AI-enhanced medical research is moving toward "agentic" AI, meaning systems that don't respond to only one prompt but instead carry out multistep research workflows with increasing autonomy.

A robot android using a typewriter

How might AI affect the workflow of biomedical research? (Image credit: Getty Images/Moor Studio)

Instead of responding with only text, an agentic agent is capable of checking and iterating on its own work until it reaches its objective. It can also take action on a user’s behalf, like searching the internet and running code, rather than just writing it.

That shift toward greater AI autonomy and less human oversight brings both enormous potential and serious risk. In a January study published in the journal Nature Biomedical Engineering, researchers evaluated LLMs on 293 coding tasks drawn from 39 published biomedical studies, initially allowing the LLMs to come up with workflows on their own. They found that the overall accuracy came in below 40%.

Their solution was to separate planning from execution: They had the AI produce a step-by-step analysis plan that a human researcher reviewed before any code got written. The approach boosted the accuracy to 74%.

The goal of AI is not perfection, but to do better than people.

Ian McCulloh, professor of computer science at Johns Hopkins University's Whiting School of Engineering

"The goal is not to ask researchers to blindly trust an AI system," study co-author Zifeng Wang, who was a doctoral student at the University of Illinois Urbana-Champaign at the time of the study, told Live Science in an email.

Instead, the goal is to "design frameworks where the reasoning, planning, and intermediate steps are visible enough that researchers can supervise and validate the process," said Wang, who is a co-founder of Keiji AI.

Why safeguards matter

These risks don't mean researchers should shy away from AI, but they do need to apply the same rigor to AI-generated work that they would to any other collaborator's output, scientists caution.

"The question is not whether LLMs accelerate science or create 'AI slop,'" Ian McCulloh, a professor of computer science at Johns Hopkins University's Whiting School of Engineering, told Live Science in an email. "The question is how we leverage this powerful technology within the scientific method."

But McCulloh also cautioned against holding AI to an impossible standard. People tend to assume AI is error-prone and downplay human error, he said, when, in reality, both humans and machines make mistakes. He anecdotally described a consulting client who lamented AI's 15% miss rate on a certain task, not realizing his human employees' miss rate was 25%.

"The goal of AI is not perfection," McCulloh said, "but to do better than people."

That effort will involve agreeing on how to measure AI's success. Dr. Ethan Goh, a physician-researcher at Stanford University, pointed out that health care still lacks standardized benchmarks for evaluating AI's performance. Goh recently published a randomized trial in JAMA Network Open that studied how LLMs influence doctors' reasoning in determining diagnoses.

RELATED STORIES

Because LLMs are trained on such a vast amount of data, "benchmarks are so expensive to produce," Goh told Live Science. What's more, he said, AI improves so quickly that most commercial models start beating the few benchmarks that exist and rapidly render them useless. Amid these challenges, Goh's team at Stanford's AI Research and Science Evaluation (ARISE) Healthcare Network is working to develop such standards by the end of this year.

For all the uncertainty around standards and safeguards, the researchers who spoke with Live Science shared a common conviction: AI belongs in the lab, but not unsupervised.

"We have to be careful not to forget what we know in terms of the scientific process," Sirota said. "But I think the opportunity is tremendous."

Patrick Sullivan has been a professional writer and editor since 2009 and producing health care content since 2015. Based in New Jersey, he is a father of two children and servant to an ever-changing number of pet rabbits. When he's not at his writing desk, you can usually find him on a yoga mat, a Brazilian jiu jitsu mat, or wandering through the woods.

You must confirm your public display name before commenting

Please logout and then login again, you will then be prompted to enter your display name.