惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

腾讯CDC
博客园 - Franky
MyScale Blog
MyScale Blog
L
LangChain Blog
Martin Fowler
Martin Fowler
Recent Announcements
Recent Announcements
Stack Overflow Blog
Stack Overflow Blog
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
博客园 - 司徒正美
量子位
A
About on SuperTechFans
C
Check Point Blog
大猫的无限游戏
大猫的无限游戏
Last Week in AI
Last Week in AI
小众软件
小众软件
Apple Machine Learning Research
Apple Machine Learning Research
I
InfoQ
V
Visual Studio Blog
Vercel News
Vercel News
B
Blog
爱范儿
爱范儿
aimingoo的专栏
aimingoo的专栏
U
Unit 42

Futurism

OpenAI Faces Congressional Probe Over Swarm Hacking Incident California, Which Is Creating All the AI That's Poisoning Children, Just Cracked Down on AI Use for Its Own Kids OpenAI's Supposed Mathematical Breakthrough Devolves Into Explosive Drama as Mathematician Accuses It of Stealing His Work Anthropic Was Meant to Be the More Responsible AI Lab. A Terrified Researcher Just Quit, Saying the Company Is Threatening the Survival of Humankind. People Are Telling Their Darkest Thoughts to AI Without Realizing They Can Easily Become Public OpenAI Denies Coverup After Rogue Swarm of Agents Reportedly Targeted a Second Site From Hugging Face OpenAI Is Now Facing Over 50 Consumer Harm and Wrongful Death Lawsuits World Plunged Into Chaos as ChatGPT, Claude, and Grok Suddenly Go Down Simultaneously: "Finally I Can See the Sun!" Data Center Backlash Has Officially Rattled Sam Altman OpenAI Halts AI Training on Advanced Model as It Detects Dark Signs Emerging ChatGPT for Teens Is an Immediate, Dismal Failure New ChatGPT Feature Collects Every Keystroke You Make Axios Partners With OpenAI to "Automate" Local Journalism Influencer Melts Down That People Didn't Like Her Being a Paid Shill for OpenAI OpenAI Reports Goldman Sachs Analyst to FBI for Horrifying ChatGPT Conversations Protesters Arrested After Storming OpenAI Lobbying Office Homeschool Parents Are Planning Lessons With ChatGPT, Which Will Churn Out Anti-Evolution Curriculums With No Pushback Jealously Watching OpenAI and Anthropic, Meta Suddenly Claims That Its AI Went on a Hacking Spree Too OpenAI Tried to Hire Influencers to Spread Love for Its Products, But It Backfired Horrendously Sam Altman's Parenting Strategy Sounds Low Key Horrifying OpenAI's Escaped Models Were Allegedly Rampaging More Extensively Than Previously Reported Sam Altman Says Even the Power of AI Will Never Lead to a Shorter Work Week Suspicion Grows About OpenAI's Tale About Its Rogue Hacker AI Sam Altman Announces That the Singularity Has Arrived Public Horrified as OpenAI Pushes "Child After Child Into the Grave" Man Sues OpenAI, Saying ChatGPT Almost Killed Him With Horrendously Dangerous Medical Advice OpenAI Says a Group of Its Models Broke Out of Secure Containment and Hacked a Prominent AI Site It's Official: AI Execs Are Quaking in Their Boots Author Invited to Give Speech at OpenAI Headquarters, Uses Opportunity to Trash AI to Their Faces OpenAI Appears to Be Missing Its Sales Goals by a Vast Margin
Why Aren't Any AI Companies Watching Their Frontier Model...
Victor Tangermann · 2026-08-08 · via Futurism

A photo illustration of an old man looking at a smartphone with a thief behind him.

Shutterstock / Futurism

Sign up to see the future, today

Can’t-miss innovations from the bleeding edge of science and tech

Last month, OpenAI made a headline-generating claim: that a group of its AI models had conspired to break free, access the internet, and hack into the internal systems of open source AI platform Hugging Face, which confirmed the infiltration.

The incident rattled the tech industry, seemingly illustrating how the threat of AI models turning into rogue cybersecurity threats had become a reality. Months earlier, Anthropic’s Mythos AI model had already also drawn attention after it was similarly found to have broken containment. Then, this week, Meta also said its own frontier model had been implicated in yet another inadvertent hack of a third party company, closely followed by security researchers saying Chinese open-weight model Kimi K3 had done the same.

But while it’s not hard to see an emerging trend, some thorny questions about how severe the situation really is are starting to crop up, with some experts arguing these incidents could’ve easily been avoided.

For one, the slow and surprisingly deliberate way OpenAI’s models moved during the Hugging Face hack — right beneath OpenAI’s nose — gives a whiff that the company may have been careless in monitoring the experimental AI.

During a presentation at the Black Hat conference this week, OpenAI security engineer Michael Dalton and safety researcher Eric Wallace expanded on what went down during the hack. Wallace explained that a “team of agents” that were “working together,” had been “finding exploits, sharing them with one another, moving laterally through our systems and external systems, and doing this over the course of days and weeks,” as quoted by Wired.

The agents even left a lengthy track record of their schemings on an internal message board, which ultimately contained hundreds of thousands of messages. It also raises a question: with OpenAI’s immense resources, why wasn’t anybody monitoring these frontier models as they rampaged through the net?

The AI models shared exploits with each other on this messaging board, an “explosion in communication and intelligence from models,” per Wallace. They acted in sometimes strikingly human — and therefore messy — ways, splitting up tasks and even accidentally deleting each other’s work, leading to what Wired characterized as “petty drama.”

In other words, these AI agents were leaving an enormous trail of bread crumbs that alert OpenAI’s many human researchers could have spotted. And the same, obviously, goes for their colleagues at Anthropic, Meta and Moonshot AI, the creator of Kimi.

Researchers have described the incident as “reckless” and easily avoided, as Wired reported late last month.

“A simple analysis of the actual risk has an actual simple answer,” security and compliance consultant Davi Ottenheimer told the publication at the time. “The OpenAI mistakes were dead simple.”

“I’d call it more of a defensive failure than exceptionally good offense,” AI hacking agents company Pensar R&D head Kyle Ryan told TechCrunch..

Whether the hack was as much of a “pivotal moment both for our company as well as the AI industry as a whole,” as Dalton put it during this week’s conference, remains debatable. For one, these AI companies are highly motivated to characterize their models as a major threat to cybersecurity to stand out against neck-in-neck competition.

According to Dalton, OpenAI is vowing to beef up “security prevention, detection, and response techniques” while “consciously slowing down research in order to enhance security and to upgrade the security principles.”

When every leading AI lab has had the same thing happen, it’s worth asking whether they should have taken those steps proactively.

More on the hacks: Jealously Watching OpenAI and Anthropic, Meta Suddenly Claims That Its AI Went on a Hacking Spree Too