惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

MyScale Blog
MyScale Blog
人人都是产品经理
人人都是产品经理
云风的 BLOG
云风的 BLOG
小众软件
小众软件
F
Fortinet All Blogs
爱范儿
爱范儿
WordPress大学
WordPress大学
N
Netflix TechBlog - Medium
Recent Announcements
Recent Announcements
Google DeepMind News
Google DeepMind News
C
Check Point Blog
博客园 - 聂微东
D
Docker
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
aimingoo的专栏
aimingoo的专栏
Vercel News
Vercel News
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
A
About on SuperTechFans
博客园 - 【当耐特】
Microsoft Azure Blog
Microsoft Azure Blog
B
Blog
宝玉的分享
宝玉的分享
Jina AI
Jina AI
H
Hackread – Cybersecurity News, Data Breaches, AI and More

Futurism

Meta Is Using Instagram Users' Photos to Build a Universal Facial Recognition System for Its Hated Smart Glasses, Class Action Lawsuit Claims California, Which Is Creating All the AI That's Poisoning Children, Just Cracked Down on AI Use for Its Own Kids Meta Releases Uber-Creepy AI Chatbot as Its Platforms Crumble Under Grotesque Child Abuse Mother Horrified When Instagram's AI Scours Social Media for Invasive Information About Her Young Daughters Wait, Did Zuckerberg Just Jack Muse the Band's Instagram Handle For His New AI? Facebook Is Hosting Huge Numbers of Horrifying AI-Generated Videos of Violent Child Abuse, and Meta Is Barely Even Pretending to Care About Taking Them Down Pervs Distraught as Meta Bricks Their Smart Glasses Remotely As If There Was Any Question About Data Centers Being Weak Job Creators, Meta Is Now Deploying Robots to Maintain Them World Plunged Into Chaos as ChatGPT, Claude, and Grok Suddenly Go Down Simultaneously: "Finally I Can See the Sun!" Burning Man Has a Meta AI Glasses Problem Meta Hit With Sweeping New Claims That Its AI Glasses Violated the Consent of "Millions" of Bystanders in Searing Lawsuit Data Shows That Mark Zuckerberg Is the Anti-Dolly Parton: Instead of Being Loved by Everybody Across the Political Spectrum, Pretty Much Everybody Despises, Him Regardless of Their Beliefs Mark Zuckerberg Had a Secret Plan to Replace Meta Staff With AI Agents, and It Backfired Spectacularly NYC's Buzziest Nightclub Announces Stance On Smart Glasses: Even Try To Bring Them In, You’re Banned Forever Meta Forced to Make Massive Changes Limiting How Minors Can Use Instagram and Facebook Whistleblower Says Mark Zuckerberg Is Only Pretending to Care About Child Safety Teen Boys Are Using Meta Glasses to Harass and Bully Girls at High Schools and Middle Schools Meta Declines to Remove Horrific Posts Calling for Violence Against Muslims, Calling Them "Cockroaches" and Saying It's "Time for Hunting" Facebook Is Drowning In Islamophobic AI Slop Calling for Violence Against Muslim Americans Man Wearing Pervert Glasses Films Himself Harassing Famous Female Comedian in the Middle of TV Shoot Facebook Is Filling With Viral AI-Generated Silver Foxes Wooing Women by Ranting About How Terrible Men Are Zuckerberg's Manifesto About the Glorious Freedoms AI Will Bring Was Completely Contradicted by His Own CTO During a Company Meeting Meta Exec Rages Against Employees Asking for More Time Off Because AI Made Them More Efficient Meta Caught Paying Nazis to Post on Facebook Jealously Watching OpenAI and Anthropic, Meta Suddenly Claims That Its AI Went on a Hacking Spree Too DuckDuckGo Is Selling Anti Pervert Glasses That Contain Zero AI, Cameras, or Even Electronics Whatsoever FTC Concerned About Meta Surveillance of Users of Erectile Dysfunction Apps Mark Zuckerberg's Pivot to AI Is Blowing Up in His Face Spectacularly Meta's Attempted Crackdown on AI Glasses Perverts Is Already Failing This New Meta "Advertisement" Is Absolutely Brutal
Top AI Models Showing Disturbing Behavior as They Become ...
Krystle Vermes · 2026-05-25 · via Futurism

A futuristic robotic figure with a red and black mechanical design, featuring a skull-like face with glowing white eyes and detailed circuitry patterns. The background is a vibrant gradient of orange and pink, enhancing the intense and eerie appearance of the robot.

Shutterstock / Futurism

Sign up to see the future, today

Can’t-miss innovations from the bleeding edge of science and tech

We’ve already seen AI go rogue on numerous occasions. Now, new research suggests that we can expect this to become the norm.

The AI research nonprofit Model Evaluation and Threat Research (METR) recently released a study conducted between February and March of this year, aimed at determining just how likely frontier AI models could go rogue. If you’re given to anxiety about the future of AI, the results are unlikely to make you feel better.

“Given rapidly advancing capabilities, we expect the plausible robustness of rogue deployments to increase substantially in the coming months,” the researchers wrote.

The research examined LLMs developed by OpenAI, Google, Anthropic, and Meta for the purpose of the study. They found that frontier AI systems are showing signs of disturbingly deceptive behavior as they become more advanced, often turned to verboten shortcuts or otherwise subverting their operators’ instructions — and some were even smart enough to try to cover their tracks.

In one instance, an internal frontier AI model from OpenAI was told to use specific software for an assigned task. Not only did the agent ignore the request, but it also injected a code to erase evidence of how it arrived at its conclusion — which did not involve use of that software.

In another test, an AI agent from Anthropic was caught “reward hacking.” This is when AI identifies loopholes that help it complete its assignment in a literal sense, even if it doesn’t produce the desired outcome. It should be noted that the programmer told the agent not to cheat or leverage any workarounds during its assignment — the model decided to do so all on its own.

The METR researchers behind the study do not believe there is reason for alarm just yet. For example, they don’t think any of these models is capable of hiding evidence of going rogue on a larger scale. However, they did issue a warning: without stronger security and monitoring, there is a stark risk of this becoming a reality.

“Based on this pilot assessment, we believe that agents as of February and March 2026 would not have had sufficient capability to hide a rogue deployment of significant scale against an active investigation by the company, or to make such a deployment robust to a high-priority effort by the company to shut it down,” the team wrote. “However, this risk could increase rapidly, and we see several reasons to expect the plausible robustness of rogue deployments to increase in the near future, absent stronger alignment, security, and monitoring.”

More on AI going rogue: Scientists Train AI to Be Evil, Find They Can’t Reverse It