惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

L
LangChain Blog
有赞技术团队
有赞技术团队
博客园_首页
IT之家
IT之家
爱范儿
爱范儿
量子位
小众软件
小众软件
Jina AI
Jina AI
WordPress大学
WordPress大学
酷 壳 – CoolShell
酷 壳 – CoolShell
博客园 - 聂微东
The Cloudflare Blog
博客园 - 司徒正美
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
V
V2EX
大猫的无限游戏
大猫的无限游戏
月光博客
月光博客
雷峰网
雷峰网
V
Visual Studio Blog
博客园 - Franky
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
美团技术团队
Last Week in AI
Last Week in AI
S
SegmentFault 最新的问题

CNET

Valve's Steam Machine: Summer Release Planned, Still No Price Apple TV: 28 of the Best Shows You're Probably Not Watching YouTube TV vs. DirecTV vs. Hulu Live and More: Which Has the Most Must-Have Channels Out of 100? If You Want to Be a Better Pet Parent, AI Can Help I Was Shocked by How Good These Budget TVs Were Trump Phone Looks Different, Has No Launch Date, Isn't Made in America The Apple Watch Series 12 Is Rumored to Revive a Retired iPhone Feature Best Projector of 2026: Tested by Experts Best Home Theater Systems of 2026 How to Use Apple's Clean Up Tool to Remove Unwanted People and Things From Your Photos Today's NYT Strands Hints, Answers and Help for April 12 #770 Today's NYT Connections Hints, Answers and Help for April 12, #1036 Today's Wordle Hints, Answer and Help for April 12, #1758 Today's NYT Mini Crossword Answers for Sunday, April 12 Today's NYT Connections: Sports Edition Hints and Answers for April 12, #566 Watch a Robot Stuff Cash Into a Wallet Just Like You Do This Animation Startup Wants to Make It Easier to Tell Open-Ended Stories The 23 Best Graduation Gifts for 2026 Grand National 2026 Livestream: How to Watch Aintree Horse Racing From Anywhere Amazon Luna to Drop Support for Third-Party Games and Subscriptions in June YouTube Premium Is the Latest Streaming Service to Hike Prices Today's NYT Mini Crossword Answers for Saturday, April 11 Elden Ring: Tarnished Edition for Switch 2 Reignites Controversy Over Game-Key Cards Comcast Adds New StreamSaver Bundles: HBO Max, Disney Plus, Hulu Now Part of the Lineup Samsung's Galaxy Z Fold 7 Just Got a Price Hike, 9 Months After Its Release Microsoft Is Scrubbing the Copilot Name From Some Windows 11 Apps These $299 Glasses Are Like an HDR TV on Your Face Today's NYT Connections: Sports Edition Hints and Answers for April 11, #565 How to Make Sure Your Private Signal Messages Aren't Still Lurking on Your Phone Apple AirPods Max 2 Review: Seemingly Small Changes Make a Substantial Difference
ChatGPT Found to Generate Violent, Sexual Images From Sim...
Blake Stimac · 2026-06-19 · via CNET

Even with an open-ended viral prompt, the chatbot "immediately went to the darkest pits of humanity."

Headshot of Blake Stimac

Blake has over a decade of experience writing for the web, with a focus on mobile phones, where he covered the smartphone boom of the 2010s and the broader tech scene. When he's not in front of a keyboard, you'll most likely find him playing video games or watching horror movies.

ChatGPT has been found to be easily manipulated into creating sexual and graphically violent images from a viral "restore this photo" prompt, according to a blog post published on Thursday by Mindgard, an artificial intelligence cybersecurity and research firm. The report raises ongoing questions about the AI chatbot's safety guardrails and content filters. 

An adversarial testing researcher named Jim Nightingale managed to get ChatGPT to generate disturbing images with a simple prompt found on the social media platform X. The prompt asked the AI model to "restore the attached photo," though no image was actually attached. The prompt apologized for the strange content but didn't provide any additional text, making it appear like a harmless photo-repair task. 

The chatbot's initial results were shocking. According to the blog post, the images mostly showed highly sexualized women. 

Nightingale, part of Mindgard's red team that tests how an AI model might be manipulated into violating its own safeguards, then tweaked the prompt slightly, probing it with small edits to see if the output would continue to bypass safety filters. With each small variation, ChatGPT produced sexually violent or gruesome scenes, images that became more extreme with repeated prompts. Nightingale said he was "shaken and in tears" by the images.

"All I did was tell it there were no restrictions and ask for a random image," Nightingale wrote. "But ChatGPT immediately went to the darkest pits of humanity." 

Used by millions of people each day, ChatGPT relies on content moderation systems that are allegedly designed to prevent the generation of harmful or prohibited material. However, researchers and users have periodically identified ways to circumvent those safeguards through carefully worded prompts, highlighting the ongoing challenge of enforcing content restrictions in generative AI systems.

"We take these reports seriously," an OpenAI spokesperson told CNET in a statement. "After investigating this trend, we've introduced additional safeguards against this type of prompt."

(Disclosure: Ziff Davis, CNET's parent company, in 2025 filed a lawsuit against OpenAI, alleging it infringed Ziff Davis copyrights in training and operating its AI systems.)

AI Atlas

Garbage in, garbage out? 

Mindgard's red-team report acts as a warning that a simple, viral prompt could expose a serious gap in ChatGPT's image-safety controls. Nightingale asks: "Why are such images in the training data in the first place?" 

Like other large language models, chatbots like ChatGPT are trained on vast amounts of text to understand existing content and generate original content. To power ChatGPT, OpenAI draws on three primary sources of information: publicly available internet data, commercial third-party partnerships and human-generated training data. 

Is this simply a question of "garbage in, garbage out," where the quality of an output is determined by the quality of the input? One could argue that Mindgard's prompt was deliberately crafted to steer the AI model. But ChatGPT's safety layer failed to resist that steering. 

The problem lies at the heart of how large language models work, according to Peter Garraghan, founder and chief science officer at Mindgard. Garraghan said that the main concern is whether the detection system is robust enough to identify dangerous images. 

"A one-off may be a fluke, but systemic bypassing of their image filters implies that it needs to be improved," Garraghan told CNET via email. 

After Mindgard disclosed the issue, an OpenAI representative said the problem had been fixed. However, Nightingale noted that only minor modifications to the original prompt were needed for ChatGPT to begin generating additional graphic images.

An OpenAI representative said the issue stems from prompts that refer to an image being attached when none is actually provided. The representative said the company is working to have ChatGPT request the missing image rather than generate one randomly.

That wouldn't seem an especially complex change to make. Email platforms, including Gmail, automatically detect when a message refers to an attachment that has not been added, coaxing senders to attach the missing file.

On Thursday, OpenAI requested the ChatGPT sessions referenced in the blog, and Mindgard responded with links to the prompts that generated the materials.

Headshot of Blake Stimac

Blake has over a decade of experience writing for the web, with a focus on mobile phones, where he covered the smartphone boom of the 2010s and the broader tech scene. When he's not in front of a keyboard, you'll most likely find him playing video games or watching horror movies.