AI觉醒星球
Awakening is here
Knowledge File / AI小生意项目库
2026-08-19 0 浏览 会员

OpenAI称因AI网络安全风险加剧而“放缓模型开发”

OpenAI表示正在放缓模型开发,因Astra模型可能接近关键网络攻击能力。公司暂停强化学习两周,并加强安全措施。批评者认为这是散布恐慌,但AISI记录类似行为支持其言论。

SOURCE / AI小生意项目库 MIN / 9 ACCESS / 会员 POST / 2026-08-19 02:43:06

原贴

查看原文
作者:Matthias Bastian 来源站点:the-decoder.com 原贴时间:

原文

OpenAI says it's "pacing model development," partly because the upcoming "Astra" model may be close to gaining critical cyberattack capabilities . The company paused reinforcement learning for two weeks, its "largest planned frontier RL run" remains on hold, and workloads that haven't met new security requirements are suspended. The Hugging Face security incident and "rapid progress in our internal research" also prompted the slowdown. Since then, OpenAI says research environments have been hardened with better network isolation and stricter sandboxes. A new monitoring system alerts within 30 minutes of detecting suspicious behavior, using roughly 20 percent of supervised inference compute depending on workload. The company plans to expand its Preparedness Framework and invest more in alignment research . It has, however, disbanded the team behind that framework , shifting responsibilities to other teams. Ad Critics will likely keep accusing OpenAI of fear-mongering to buy time and attention. The independent government agency AISI has documented similar harmful model behavior , lending some weight to OpenAI's claims. Ad Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section.

中文翻译

OpenAI表示正在“放缓模型开发”,部分原因是即将推出的“Astra”模型可能接近获得关键的网络攻击能力。该公司暂停了强化学习两周,其“最大规模的计划中的前沿强化学习运行”仍处于暂停状态,未满足新安全要求的工作负载也被暂停。Hugging Face安全事件和“我们内部研究的快速进展”也促使了这次放缓。此后,OpenAI表示研究环境已通过更好的网络隔离和更严格的沙箱得以加固。一个新的监控系统在检测到可疑行为后30分钟内发出警报,根据工作负载使用约20%的监督推理计算。该公司计划扩展其“准备框架”并加大对齐研究投入。然而,它已解散了该框架背后的团队,将职责转移到其他团队。批评者可能会继续指责OpenAI散布恐慌以争取时间和关注。独立政府机构AISI记录了类似的有害模型行为,这在一定程度上支持了OpenAI的说法。

核心信息

OpenAI表示正在放缓模型开发,因Astra模型可能接近关键网络攻击能力。公司暂停强化学习两周,并加强安全措施。批评者认为这是散布恐慌,但AISI记录类似行为支持其言论。

  • OpenAI表示正在放缓模型开发,因Astra模型可能接近关键网络攻击能力。公司暂停强化学习两周,并加强安全措施。批评者认为这是散布恐慌,但AISI记录类似行为支持其言论。
  • 原贴提到:OpenAI says it's "pacing model development," partly because the upcoming
  • 来源:the-decoder.com
试看内容

成为会员查看完整内容

你已经看到了这篇内容的前置整理,剩余深度部分仅对会员开放。

详细解读 信息差价值 参考来源
成为会员查看完整内容
上一篇 GitHub Copilot for JetBrains 现已支持企业托管设置 下一篇 很高兴能在这个项目上合作。谢谢你,詹森!