惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

博客园 - 叶小钗
D
Docker
Google DeepMind News
Google DeepMind News
Y
Y Combinator Blog
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
Blog — PlanetScale
Blog — PlanetScale
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
U
Unit 42
博客园 - 【当耐特】
N
Netflix TechBlog - Medium
V
Visual Studio Blog
Microsoft Azure Blog
Microsoft Azure Blog
博客园_首页
Recent Announcements
Recent Announcements
GbyAI
GbyAI
T
Tailwind CSS Blog
S
SegmentFault 最新的问题
WordPress大学
WordPress大学
T
The Blog of Author Tim Ferriss
Engineering at Meta
Engineering at Meta
L
LangChain Blog
A
About on SuperTechFans
M
MIT News - Artificial intelligence
B
Blog

OpenAI News

Using custom GPTs ChatGPT for customer success teams Applications of AI at OpenAI Research with ChatGPT Analyzing data with ChatGPT Financial services Responsible and safe use of AI Writing with ChatGPT ChatGPT for research Creating images with ChatGPT Personalizing ChatGPT ChatGPT for finance teams Getting started with ChatGPT Working with files in ChatGPT Learn ChatGPT workflows for sales teams Prompting fundamentals ChatGPT for managers Using projects in ChatGPT Learn ChatGPT workflows for marketing teams Brainstorming with ChatGPT AI fundamentals ChatGPT for operations teams Healthcare Our response to the Axios developer tool compromise Using skills OpenAI Full Fan Mode Contest: Terms & Conditions CyberAgent moves faster with ChatGPT Enterprise and Codex The next phase of enterprise AI 儿童安全蓝图正式发布 推出 OpenAI 安全研究员计划
打击恶意使用 AI 的行为:2025 年 10 月
2025-10-07 · via OpenAI News

我们的使命是确保通用人工智能造福全人类。我们推进这一使命的方式就是部署创新技术,让这些技术帮助人们解决棘手问题,同时构建基于常识规则的民主人工智能,以此保护人们免受实际伤害。

2024 年 2 月开始发布威胁报告以来,我们已打击并报告了 40 多个违反我们使用政策的网络。其中包括防止威权政权利用 AI 控制民众或胁迫他国,以及防范欺诈、恶意网络活动和秘密影响行动等滥用行为。

在本期报告中,我们分享了过去一个季度的案例研究,以及我们如何侦测和打击恶意使用模型的行为。我们观察到,威胁行为者继续将 AI 嫁接到旧有攻击脚本上以加快行动速度,而不是利用我们的模型获取新型攻击能力。如果发现违反政策的活动,我们就会封禁相关帐户,并视需要与合作伙伴共享相关洞察。通过公开报告、政策执行和同行协作,我们致力于提高公众对滥用行为的认知,同时加大对普通用户的保护力度。