惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

G
Google Developers Blog
F
Fortinet All Blogs
Microsoft Azure Blog
Microsoft Azure Blog
腾讯CDC
Vercel News
Vercel News
Recent Announcements
Recent Announcements
博客园 - Franky
小众软件
小众软件
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
The Cloudflare Blog
宝玉的分享
宝玉的分享
I
InfoQ
博客园 - 聂微东
Jina AI
Jina AI
J
Java Code Geeks
V
V2EX
U
Unit 42
Stack Overflow Blog
Stack Overflow Blog
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
阮一峰的网络日志
阮一峰的网络日志
L
LangChain Blog
T
The Blog of Author Tim Ferriss
量子位

Latest news

LG G6 vs. LG G5: I compared the latest OLED TV models, and it's a surprisingly tough choice I saw the 'MacBook Pro for Linux users' for the first time, and it's a legit Windows threat I'm putting Motorola above Samsung when it comes to flip phones - and won't think twice I got an early look at ChatGPT Images 2.0, and it's impressive - with one exception I compared Thread, Zigbee, and Matter - here's the best smart home setup for you Scaling agentic AI demands a strong data foundation - 4 steps to take first 5 Apple products explain my optimism for John Ternus as the next CEO This Motorola phone deal comes with free Bluetooth trackers and earbuds - how it works Can a near infrared laptop light boost my mood? I tried it to find out This free app makes journaling so easy that I've managed to do it for 3 months CachyOS is the Arch Linux distro to try if you want serious speed and performance I tested Surfshark's new Dausos VPN protocol - here's how it compares to WireGuard How to easily encrypt your files on an Android phone - for free I'm not giving up on DJI cameras yet - not when they can upset my GoPro like this Why I'm recommending last year's phones over 2026 models - with one exception This powerful Gemini setting made my AI results way more personal and accurate After testing this HP laptop, I get why its 'boring' design is adored by business users The best TV antenna of 2026: Expert tested Your old iPad or Android tablet can be your new smart home panel - here's how Apple's original AirTag still tracks effectively, and you can get a 4-pack for its best price ever T-Mobile will give you an iPad for $99 when you sign up for a new line - here's how How to qualify for Apple's education discount - and get a $499 MacBook Neo for school T-Mobile will give you a Samsung Galaxy Watch 8 for free - how to get yours Verizon will give you a free iPad or Apple Watch with your next iPhone - how the deal works The best laptops of 2026: Expert tested and reviewed I hid 4 Bluetooth trackers (including AirTags) to test their reliability - here's how Android rivals compared I stopped using my iPhone's hotspot after testing this 5G router - and that won't change The best Kindles in 2026: Expert recommended Does Best Buy price match? Everything to know about matching prices online and in-store The best WordPress hosting services of 2026: Expert tested and reviewed
Prolonged AI use can be hazardous to your health and work...
2026-04-18 · via Latest news
adeephole-gettyimages-856656258
Puneet Vikram Singh, Nature and Concept photographer/ Moment via Getty Images

Follow ZDNET: Add us as a preferred source on Google.


ZDNET's key takeaways 


  • AI is getting better at small tasks, but still lags on long-form analysis. 
  • The consequences of prolonged interactions with AI can be disastrous.
  • Use AI like a tool for well-defined tasks, and avoid falling down a rabbit hole.

Better to do a little well than a great deal badly. So said the great philosopher Socrates, and his advice can apply to your use of artificial intelligence, including chatbots such as OpenAI's ChatGPT, or Perplexity, as well as the agentic AI programs increasingly being tested in enterprise.

AI research increasingly shows that the safest and most productive course with AI is to use it for small, limited tasks, where outcomes can be well defined, and results can be verified, rather than pursuing extensive interactions with the technology over hours, days, and weeks.

Also: Asking AI for medical advice? There's a right and wrong way, one doctor explains

Extended interactions with chatbots such as ChatGPT and Perplexity can lead to misinformation at the very least, and in some cases, delusion and death. The technology is not yet ready to take on the most sophisticated kinds of demands of reasoning, logic, common sense, and deep analysis -- areas where the human mind reigns supreme. 

(Disclosure: Ziff Davis, ZDNET's parent company, filed an April 2025 lawsuit against OpenAI, alleging it infringed Ziff Davis copyrights in training and operating its AI systems.)

We are not yet at AGI (Artificial General Intelligence), the supposed human-level capabilities of AI, so you'd do well to keep the technology's limitations in mind when using it.

Put simply, use AI as a tool rather than letting yourself be sucked down a rabbit hole and get lost in endless rounds of AI conversation.

What AI does well - and not so well

AI tends to do well at simple tasks, but poorly at complex and deep types of analysis.

The latest examples of that are the main takeaways from this week's release of the Annual AI Index 2026 from Stanford University's Human-Centered AI group of scholars. 

On the one hand, editor-in-chief Sha Sajadieh and her collaborators make clear that agentic AI is increasingly successful at tasks such as looking up information on the Web. In fact, agents are close to human-level on routine online processes.

Also: 10 ways AI can inflict unprecedented damage 

Across three benchmark tests -- GAIA, OSWorld, and WebArena -- Sajadieh and team found that agents are approaching human-level performance on multi-step tasks such as opening a database, applying a policy rule, and then updating a customer record. On the GAIA test, agents have an accuracy rate of 74.5%, still below the 92% of human performance but way up from the 20% of a year ago.

On the OSWorld test, "Computer science students solve about 72% of these tasks with a median time of roughly two minutes," while Anthropic's Claude Opus 4.5, up until recently its most powerful model, reaches 66.3%. That means "the best model [is] within 6 percentage points of human performance."

WebArena shows AI models "now within 4 percentage points of the human baseline of 78.2%" accuracy.

stanford-hai-agentic-ai-on-gaia-test

Agentic AI is getting better at online tasks such as Web browsing but still falls short of human-level accuracy.

Stanford

While Claude Opus and other LLMs are not perfect, they show rapid progress in at least reaching benchmark levels that come closer to human-level performance.

That makes sense, as manipulating a web browser or looking something up in a database should be among the easier scenarios in which the natural-language prompt can plug into APIs and external resources. In other words, AI should have most of the equipment it requires to interface with applications in limited ways and carry out tasks.

Also: 40 million people globally are using ChatGPT for healthcare - but is it safe?

Note that even with well-defined, limited tasks, it helps to check what you're getting from a bot, as the average score on these benchmarks still falls short of human capacity -- and that's in benchmark tests, a kind of simulated performance. In real-world settings, your results may vary, and not to the upside.

AI can't handle the hard stuff

When they dug into deeper kinds of work, the Stanford scholars found much less encouraging results.

Research has found, they noted, that "models handle simple lookups well but struggle when asked to find multiple pieces of matching information or to apply conditions across a very long document -- tasks that would be straightforward for a human scanning the same text."

That finding aligns with my own anecdotal experience using ChatGPT to draft a business plan. Answers were fine in the first few rounds of prompting, but then degraded as the model snuck in facts and figures I had not specified, or that might have been relevant earlier in the process but had no business being included in the present context.

The lesson, I concluded, was that the longer your ChatGPT sessions, the more errors sneak in. It makes the experience infuriating.

Also: I built a business plan with ChatGPT and it turned into a cautionary tale

The results of unchecked bot elaboration can get more serious. An article last week in Nature magazine describes how scientist Almira Osmanovic Thunström, a medical researcher at the University of Gothenburg, and her team invented a disease, "bixonimania," which they described as an eye condition resulting from excessive exposure to blue light from computer screens.

They wrote formal research papers on the made-up condition, then published them online. The papers got picked up in bot-based searches. Most of the large language models, including Google's Gemini, began to faithfully relate the condition bixonimania in chats, pointing to the faked research papers of Thunström and team. 

The fact that bots will confidently assert the existence of the fake bixonimania speaks to a lack of oversight of the technology's access to information. Without proper checking, you can't know if a model will verify what it's spitting out. As one scholar who wasn't involved in the research noted, "We should evaluate [the AI model] and have a pipeline for continuous evaluation."

Consequences can be serious

A more serious variant, where a user seems to have gone down a rabbit hole of confiding in a bot, is described in a recent New York Times article by Teddy Rosenbluth about the case of an older man grappling with white blood cell cancer. 

Rather than following his oncologist's advice, the patient, Joe Riley, relied on extensive interaction with chatbots, especially Perplexity, to refute the doctor's diagnosis. He insisted his AI research revealed he had what's called Richter's Transformation, a complication of cancer that would be made more adverse by the recommended treatment.

Also: Use Google AI Overview for health advice? It's 'really dangerous,' investigation finds

Despite emails from experts on Richter's questioning the material in the Perplexity summaries of the condition, Riley stuck with his belief in his AI-generated reports and resisted his doctor's and his family's pleas. He missed the window for proper treatment, and by the time he relented and agreed to try treatment, it was too late.

Rosenbluth makes the connection between the story of Joe Riley and the case of Adam Raine last year, who committed suicide after extensive chats with ChatGPT about his inclination to end his life.

Riley's son, Ben Riley, wrote his own account of his father's journey with AI. While the younger Riley doesn't blame the technology per se, he points out that getting immersed in chats and losing perspective can have consequences. 

"The fact remains that AI does exist in our world," writes Riley, "and just as it can serve as fuel to those suffering manic psychosis, so too may it affirm or amplify our mistaken understanding of what's happening to us physically and medically."

Staying sane with unreliable AI

The inclination to engage in long-form discussions about depression, suicide, and serious health conditions is understandable. People have been habituated to long-form engagements of hours at a time on social media. Some people are lonely, and a natural-language conversation with a bot is better than no conversation at all. 

Also: Your chatbot is playing a character - why Anthropic says that's dangerous

Bots have a tendency toward sycophancy, research has shown, which can make hours of engagement with a bot more fulfilling than the ordinary give and take with a person.

And the companies that make the technology, while warning users to verify bot output, have tended to place less emphasis on negative reports from individuals such as Riley and Raine.

4 rules for avoiding the rabbit hole

A few rules can help mitigate the worst effects of putting too much emphasis on the technology.

  1. Define what you are going to a chatbot for. Is there a well-defined task that has a limited scope and for which the predictions of the bot can be fact-checked with other sources?
  2. Have a healthy skepticism. It's well known that chatbots are prone to confabulation, confidently asserting falsehoods. It doesn't matter how many chatbots you use to try to balance the good and the bad; all of them should be treated with a healthy skepticism as having only part of the truth, if any.
  3. Regard chatbots not as friends or confidants. They are digital tools, like Word or Excel. You're not trying to have a relationship with a bot but rather to complete a task. 
  4. Use proven digital overload skills. Take stretch breaks. Step away from the computer for a non-digital human interaction, such as playing card games with a friend or going for a walk. 

Also: Stop saying AI hallucinates - it doesn't. And the mischaracterization is dangerous

Falling down the rabbit hole happens partly as a result of simply being parked in front of a screen with no downtime.