惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

V
V2EX
P
Proofpoint News Feed
D
DataBreaches.Net
C
Check Point Blog
L
LangChain Blog
量子位
美团技术团队
Vercel News
Vercel News
人人都是产品经理
人人都是产品经理
N
Netflix TechBlog - Medium
V
Visual Studio Blog
Microsoft Security Blog
Microsoft Security Blog
博客园 - 【当耐特】
MongoDB | Blog
MongoDB | Blog
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
Last Week in AI
Last Week in AI
The GitHub Blog
The GitHub Blog
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
U
Unit 42
腾讯CDC
M
MIT News - Artificial intelligence
Microsoft Azure Blog
Microsoft Azure Blog
Blog — PlanetScale
Blog — PlanetScale

Mashable

AdultFriendFinder 2016 data breach: Security improvements 5 AdultFriendFinder scams to avoid The best hookup apps of 2026: I swiped until my thumb hurt How to delete your AdultFriendFinder account Tax Day 2026 deals: Score free food from Burger King, Krispy Kreme, Popeyes, Wendy's, and more XChat to launch on iPhone and iPad The 9 best headphones and earbuds for working out in 2026 Health chatbots could pave the way for 'AI privilege' in court UFC 2026 livestream: How to watch UFC for free 'Mexodus' review: This live-looped musical is a theatrical miracle 'Zelda: Ocarina of Time' remake: 4 things I really, really want Boston Bruins vs. Tampa Bay Lightning 2026 livestream: How to watch NHL for free The DJI Mini 5 Pro drone is down to its record-low price at Amazon — save over $500 Best Hulu deals and bundles: Best streaming deals in April 2026 NYT Connections Sports Edition hints and answers for April 11: Tips to solve Connections #565 NYT Strands hints, answers for April 11, 2026 Today's Hurdle hints and answers for April 11, 2026 NYT Pips hints, answers for April 11, 2026 NYT Connections hints and answers for April 11. Tips to solve 'Connections' #1035. Wordle today: The answer and hints for April 11, 2026 Artemis 2 splashdown: Photos, videos of the astronauts' return Artemis II crew return to Earth with perfect splashdown All the streaming apps that raised prices in 2026 so far Artemis II: All the Apple, GoPro, and Microsoft gadgets on Orion 'Moon joy' takes off as NASA embraces a new space-age catchphrase The pros and cons of switching from Kindle to Kobo e-readers Apple will close its first unionized retail store 'The AI Doc' director: Cynicism is the only wrong answer to AI Artemis II return: How to livestream reentry and splashdown BTS 'Arirang' World Tour: How to watch it live in cinemas
OpenAI follows Anthropic's lead in limited release of GPT...
Amanda Yeo · 2026-04-15 · via Mashable

OpenAI has unveiled GPT-5.4-Cyber, a new AI model that may be willing to accept seemingly malicious prompts in the name of cybersecurity. Fortunately, the ChatGPT developer won't let just anyone play with its less restrictive, more freewheeling AI.

Announced via a blog post on Tuesday, GPT-5.4-Cyber is a variant of OpenAI's publicly available GPT-5.4 large language model. According to OpenAI, its frontier AI models such as GPT-5.4 have safeguards against clearly malicious use, making them refuse harmful user requests such as stealing credentials or finding vulnerabilities in code. In contrast, the company's new GPT-5.4-Cyber model is trained to be more lenient, and potentially accept these prompts instead. 

Describing GPT-5.4-Cyber as "cyber-permissive," OpenAI states that this change is to allow the AI to be used for defensive cybersecurity measures, such as helping researchers find vulnerabilities to be addressed.

"We want to empower defenders by giving broad access to frontier capabilities, including models which have been tailor-made for cybersecurity," wrote OpenAI. "This is a version of GPT‑5.4 which lowers the refusal boundary for legitimate cybersecurity work and enables new capabilities for advanced defensive workflows."

Given the potential danger posed by GPT-5.4-Cyber's lowered safeguards, not everyone will be able to immediately dive in to push the AI's arguably flexible ethical limits even further. OpenAI states that it is starting with "limited, iterative deployment to vetted security vendors, organizations, and researchers." As such, only members of its Trusted Access for Cyber⁠ (TAC) program will be given access to GPT-5.4-Cyber at present, and only those at its highest tiers. 

Mashable Light Speed

Introduced in February, TAC is a network of users who have been through OpenAI's automated identity verification process, including completing a government ID check. Once approved, users in OpenAI's TAC program are allowed access to versions of its AI models with fewer safeguards, such as GPT‑5.4‑Cyber. OpenAI states that this is intended to enable cybersecurity research, education, and programming. 

Not every TAC-approved user will immediately get their hands on GPT-5.4-Cyber, however. OpenAI states that users who aren't already part of TAC's higher tiers may request access to it, which will require going through further authentication to verify themselves as "legitimate cyber defenders." 

GPT-5.4-Cyber's reveal comes just one week after OpenAI competitor Anthropic announced Project Glasswing. Like TAC, Project Glasswing is an initiative that restricts Anthropic's cybersecurity-focused Claude Mythos Preview AI model to select approved organisations. Claiming that Claude Mythos Preview "has already found thousands of high-severity vulnerabilities," Anthropic stated that Project Glasswing was an effort to ensure its AI model was used for solely defensive cybersecurity purposes.

"Given the rate of AI progress, it will not be long before such capabilities proliferate, potentially beyond actors who are committed to deploying them safely," Anthropic wrote.


Disclosure: Ziff Davis, Mashable’s parent company, in April 2025 filed a lawsuit against OpenAI, alleging it infringed Ziff Davis copyrights in training and operating its AI systems.