惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

罗磊的独立博客
小众软件
小众软件
The Cloudflare Blog
博客园 - 【当耐特】
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
酷 壳 – CoolShell
酷 壳 – CoolShell
WordPress大学
WordPress大学
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
V
Visual Studio Blog
量子位
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
美团技术团队
S
SegmentFault 最新的问题
宝玉的分享
宝玉的分享
博客园 - 叶小钗
月光博客
月光博客
Apple Machine Learning Research
Apple Machine Learning Research
T
Tailwind CSS Blog
博客园 - 聂微东
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
J
Java Code Geeks
Y
Y Combinator Blog
D
Docker
Microsoft Azure Blog
Microsoft Azure Blog

Futurism

OpenAI Faces Congressional Probe Over Swarm Hacking Incident California, Which Is Creating All the AI That's Poisoning Children, Just Cracked Down on AI Use for Its Own Kids OpenAI's Supposed Mathematical Breakthrough Devolves Into Explosive Drama as Mathematician Accuses It of Stealing His Work Anthropic Was Meant to Be the More Responsible AI Lab. A Terrified Researcher Just Quit, Saying the Company Is Threatening the Survival of Humankind. People Are Telling Their Darkest Thoughts to AI Without Realizing They Can Easily Become Public OpenAI Denies Coverup After Rogue Swarm of Agents Reportedly Targeted a Second Site From Hugging Face OpenAI Is Now Facing Over 50 Consumer Harm and Wrongful Death Lawsuits World Plunged Into Chaos as ChatGPT, Claude, and Grok Suddenly Go Down Simultaneously: "Finally I Can See the Sun!" Data Center Backlash Has Officially Rattled Sam Altman ChatGPT for Teens Is an Immediate, Dismal Failure New ChatGPT Feature Collects Every Keystroke You Make Axios Partners With OpenAI to "Automate" Local Journalism Influencer Melts Down That People Didn't Like Her Being a Paid Shill for OpenAI OpenAI Reports Goldman Sachs Analyst to FBI for Horrifying ChatGPT Conversations Protesters Arrested After Storming OpenAI Lobbying Office Homeschool Parents Are Planning Lessons With ChatGPT, Which Will Churn Out Anti-Evolution Curriculums With No Pushback Why Aren't Any AI Companies Watching Their Frontier Models to Make Sure They Don't Go on Hacking Sprees? Jealously Watching OpenAI and Anthropic, Meta Suddenly Claims That Its AI Went on a Hacking Spree Too OpenAI Tried to Hire Influencers to Spread Love for Its Products, But It Backfired Horrendously Sam Altman's Parenting Strategy Sounds Low Key Horrifying OpenAI's Escaped Models Were Allegedly Rampaging More Extensively Than Previously Reported Sam Altman Says Even the Power of AI Will Never Lead to a Shorter Work Week Suspicion Grows About OpenAI's Tale About Its Rogue Hacker AI Sam Altman Announces That the Singularity Has Arrived Public Horrified as OpenAI Pushes "Child After Child Into the Grave" Man Sues OpenAI, Saying ChatGPT Almost Killed Him With Horrendously Dangerous Medical Advice OpenAI Says a Group of Its Models Broke Out of Secure Containment and Hacked a Prominent AI Site It's Official: AI Execs Are Quaking in Their Boots Author Invited to Give Speech at OpenAI Headquarters, Uses Opportunity to Trash AI to Their Faces OpenAI Appears to Be Missing Its Sales Goals by a Vast Margin
OpenAI Halts AI Training on Advanced Model as It Detects ...
Maggie Harrison Dupré · 2026-08-21 · via Futurism

A robotic hand typing on a keyboard with a red digital overlay showing a password input field, a padlock icon, and warning symbols, suggesting cybersecurity or hacking themes.

Sign up to see the future, today

Can’t-miss innovations from the bleeding edge of science and tech

OpenAI says that it’s slowing down development and release of new models due to security and alignment concerns.

The ChatGPT maker announced the decision in a Tuesday blog post, citing two events as drivers of the indefinite training halt. One was the recent incident in which an OpenAI agent escaped its training sandbox without OpenAI’s knowledge and coordinated with other agents to launch a bizarre cyberattack against the AI training repository Hugging Face in an effort to cheat on its training tests. The blog post also — more mysteriously — cited “preliminary evidence” that an unreleased new model called Astra “may meet the critical cybersecurity capability threshold” under OpenAI’s “Preparedness Framework,” which mandates that OpenAI slow down development if a model “could introduce unprecedented new pathways to severe harm.”

OpenAI further said that it’s in the process of rewriting its Preparedness Framework, its foundational safety document, to keep up with the emergent behaviors of “increasingly capable systems.”

“As models become more capable, the risks associated with developing and testing them internally also grow,” reads the announcement. “Our standards for monitoring, alignment, and security must stay ahead of those risks. We wanted to take the time necessary to meet those standards, so we temporarily slowed the pace of scaling.”

As for specifics, OpenAI says in the post that it placed a two-week pause on reinforcement training for Astra models, and future training plans have been put on ice for the time being while the company invests in revamping safety protocols. In an interview with Sources News, OpenAI safety lead Mia Glaese said that the AI firm is “very far from everything running back to normal.”

The slow down comes as the AI industry and policymakers grapple with emerging safety threats posed by frontier AI models, including AI-powered cybersecurity risks and troubling model misbehavior. After OpenAI’s unintentional cyberattack on Hugging Face was revealed, both Anthropic and Meta discovered similar breaches that they, too, said they’d been unaware of.

“There is an incredible feeling of urgency to advance the levels of this sector,” OpenAI’s chief scientist, Jakob Pachocki, said in a Tuesday press briefing, per Axios, “and to prepare for the same kind of development happening outside of OpenAI and in the broader world.”

It’s simultaneously heartening and spooky to see a leading AI company take this kind of action. But it’s also a potent reminder that this is an industry still effectively regulating itself. If OpenAI wants to speed back up, that’s the company’s choice to make.

More on OpenAI: New ChatGPT Feature Collects Every Keystroke You Make