惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

B
Blog RSS Feed
WordPress大学
WordPress大学
博客园_首页
罗磊的独立博客
D
Docker
N
Netflix TechBlog - Medium
博客园 - Franky
Hugging Face - Blog
Hugging Face - Blog
D
DataBreaches.Net
I
InfoQ
L
LangChain Blog
GbyAI
GbyAI
V
V2EX
博客园 - 聂微东
P
Proofpoint News Feed
博客园 - 【当耐特】
腾讯CDC
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
量子位
Martin Fowler
Martin Fowler
有赞技术团队
有赞技术团队
U
Unit 42
博客园 - 司徒正美
大猫的无限游戏
大猫的无限游戏

Latest Science News -- ScienceDaily

Researchers block key protein that helps Parkinson’s spread through the brain Scientists thought brain inflammation was driving long COVID but the scans told a different story Scientists break 30-year superconductivity record at normal pressure Tiny “sesame” sea slug discovered in Taiwan turns out to be a new species Popular anti-aging drug combo caused severe brain damage in mice New laser heat treatment could stop blindness before it starts NASA’s Webb telescope discovers a planet where rock clouds vanish every night NASA’s Fermi telescope reveals the power source behind monster supernovae Scientists say guava juice could make iron supplements work better Humanity has already exceeded Earth’s limits, study warns Scientists discover ancient single-celled ancestors still live on in your blood Scientists are raising new questions about vitamin B12 and cancer Scientists create supercharged vitamin K that helps the brain heal itself Scientists say they’ve reversed brain aging with a simple nasal spray Large Hadron Collider detects strange particle behavior that could rewrite physics AI-powered spectrometer chip shrinks lab technology to the size of a grain of sand Scientists create global treasure map pointing to hidden rare earth deposits Queenless wasp colonies explode into chaos but hidden helpers save them Deadly fungus and lung parasites are hammering wild rattlesnakes Venomous Himalayan pit viper was actually 5 different species all along NASA’s Psyche spacecraft uses Mars as a giant slingshot toward a mysterious metal world Scientists discover a giant “planet factory” beyond Jupiter Massive supercomputer simulations unlock cosmic magnetic mystery USC scientists discover a hidden Alzheimer’s trigger and a possible way to shut it down Eating more beans and soy could slash high blood pressure risk by nearly 30% Scientists discover why Ozempic and Wegovy weight loss eventually plateaus This prehistoric fish may explain how animals first walked on Earth 100-million-year-old bug had crab-like claws unlike any known insect Common heart drug taken by millions found useless — and possibly dangerous AI won’t replace you but someone using AI might
A classic brain test exposed AI's biggest weakness
2026-06-10 · via Latest Science News -- ScienceDaily

Artificial intelligence systems can write essays, answer questions, and solve complex problems. But new research suggests they may struggle with something humans do every day: staying focused on the task at hand when distractions get in the way.

Researchers led by Suketu Patel put several leading AI models through a well-known psychology experiment called the Stroop task. The results revealed a significant difference between how AI systems process information and how the human brain manages attention.

What Is the Stroop Task?

The Stroop task is a classic psychological test that has been used for decades to study attention, concentration, and self-control.

In the test, color words such as "red," "blue," or "green" are displayed in colored ink. Sometimes the word and the ink color match. For example, the word "red" might appear in red ink. Other times they conflict, such as the word "red" printed in blue ink.

Participants are asked to name the color of the ink rather than read the word itself.

That sounds simple, but it creates a challenge because reading words is an automatic habit for most people. The brain must suppress the urge to read the word and instead focus on identifying the ink color.

Psychologists often use the task to measure what is known as executive control, a set of mental processes that helps people regulate attention, resist distractions, and stay focused on goals.

Testing AI Attention

The researchers wanted to see whether modern large language models (LLMs) handle this challenge in the same way humans do.

LLMs are the AI systems behind tools such as ChatGPT, Claude, and Gemini. They are trained on enormous amounts of text and learn patterns in language, allowing them to generate responses that often appear remarkably human.

When given short lists containing five color words, the AI systems generally performed well, even when the words and colors did not match.

However, the picture changed dramatically as the lists became longer.

GPT-4o achieved 91% accuracy when working with five words. At ten words, its accuracy fell to 57%. When the list expanded to forty words, accuracy dropped to just 15%.

Claude 3.5 Sonnet maintained stable performance through lists of twenty words but then experienced a sharp decline, falling to 24% accuracy with forty-word lists.

The researchers observed similar patterns in GPT-5, Claude Opus 4.1, and Gemini 2.5.

When AI Loses Focus

The challenge became even more difficult when matching and mismatched color words appeared together in the same list.

Under those conditions, performance deteriorated further. Accuracy for the mismatched items dropped to nearly zero in some cases.

According to the researchers, the AI models had trouble maintaining the instruction to identify ink colors. Instead, they increasingly defaulted to reading the words themselves.

In other words, the systems appeared unable to consistently suppress the response they had been most heavily trained to produce.

This finding is particularly interesting because humans face a similar conflict. People are generally much better at reading words than naming ink colors. Yet despite this bias, most individuals can maintain high accuracy and stable performance even when confronted with long lists of conflicting words and colors.

Human Attention vs. Machine Attention

The study highlights an important distinction between human and artificial intelligence.

Although modern AI systems can produce impressive language and reasoning capabilities, their underlying mechanisms differ from the attention processes found in biological brains.

Humans can often sustain focus on a specific goal while filtering out competing information. The results suggest that current AI models may struggle with this type of cognitive control when tasks become increasingly demanding.

The researchers argue that the performance collapse seen in these experiments points to fundamental limitations in today's large language models. While AI can sometimes mimic human behavior, its ability to maintain attention appears to operate very differently from the way people do.

The findings offer a reminder that even the most advanced AI systems still have weaknesses, particularly when tasks require them to resist distractions and stay focused over extended sequences of information.