AI觉醒星球
Awakening is here
Knowledge File / AI小生意项目库
2026-07-25 0 浏览 会员

Opus 5 可能已解决基于浏览器的提示注入,这是困扰AI代理的最大安全漏洞

Anthropic 的 Opus 5 模型结合自动模式,在针对浏览器代理的提示注入攻击中实现了 0% 的成功率,标志着 AI 安全领域的重大突破。

SOURCE / AI小生意项目库 MIN / 9 ACCESS / 会员 POST / 2026-07-25 18:43:36

原贴

查看原文
作者:Matthias Bastian 来源站点:the-decoder.com 原贴时间:

原文

Anthropic says Opus 5 is nearly immune to prompt injections in its own software. Prompt injection , where an attacker slips past an AI model's instructions through manipulated inputs like hidden text on a webpage, fails against Opus 5 in almost every case. For browser agents, the attack success rate hit zero percent across 129 test scenarios, per the system card . That's a big deal given that OpenAI admitted in December that prompt injection may never be fully solved . In a general prompt injection test by security firm Gray Swan , the success rate after 15 attempts dropped from 5.5 percent (Opus 4.8) to 2.0 percent. That zero percent rate only holds with Auto Mode turned on in products like Claude Cowork . Auto Mode stacks two defense layers. One scans incoming data for hidden instructions before the model processes them. The other blocks dangerous actions before execution. An attacker has to beat both independently. Without them, Opus 5 sits at 3.7 percent, and Sonnet 5 actually does better at 0.93 percent. Only the combination of model and protective software pushes the rate to zero. Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section.

中文翻译

Anthropic表示,Opus 5在其自有软件中几乎能免疫提示注入。提示注入是指攻击者通过操纵输入(如网页上的隐藏文本)绕过AI模型的指令,在几乎所有情况下都无法突破Opus 5。

核心信息

Anthropic 的 Opus 5 模型结合自动模式,在针对浏览器代理的提示注入攻击中实现了 0% 的成功率,标志着 AI 安全领域的重大突破。

  • Anthropic 的 Opus 5 模型结合自动模式,在针对浏览器代理的提示注入攻击中实现了 0% 的成功率,标志着 AI 安全领域的重大突破。
  • 原贴提到:Anthropic says Opus 5 is nearly immune to prompt injections in its own s
  • 来源:the-decoder.com
试看内容

成为会员查看完整内容

你已经看到了这篇内容的前置整理,剩余深度部分仅对会员开放。

详细解读 信息差价值 参考来源
成为会员查看完整内容
上一篇 OpenClaw 签署微软开放权重与 AI 领导力信函 下一篇 Anthropic的Claude Opus 5成本远低于Fable 5,同时在大多数基准测试中与之持平或超越