AI觉醒星球
Awakening is here
Knowledge File / AI小生意项目库
2026-07-25 0 浏览 会员

Anthropic声称其新的Claude Opus 5以一半的token价格提供接近Fable 5的性能

Anthropic推出Claude Opus 5,性能接近旗舰Fable 5但价格减半,在编程和知识工作基准测试中领先,并引入可调节的effort设置以优化性价比。

SOURCE / AI小生意项目库 MIN / 9 ACCESS / 会员 POST / 2026-07-25 02:36:42

原贴

查看原文
作者:Matthias Bastian 来源站点:the-decoder.com 原贴时间:

原文

Anthropic is positioning Claude Opus 5 as a cheaper alternative to Fable 5 that beats it in several benchmarks. Opus 5 leads tests in agentic coding and knowledge work. On ARC-AGI-3, which measures novel problem-solving, the model scores 30.2 percent, nearly four times higher than GPT-5.6 Sol. Anthropic says Opus 5 can check and improve its own work through iteration, and build its own tools through code when it needs them. Anthropic's new flagship model Claude Opus 5 posts top scores in coding and knowledge work while approaching Claude Fable 5's performance at half its token rates. Anthropic is responding to pricing pressure from GPT-5.6 Sol and Chinese competitors . Its new Opus 5 model is designed to close the price-performance gap with the much pricier Fable 5 . Opus 5 becomes the default model on Claude Max and the most capable model available on Claude Pro. The 1 million-token context window and token rates remain unchanged . Like its predecessor Opus 4.8, Opus 5 costs $5 per million input tokens and $25 per million output tokens. A new Fast Mode increases speed by 2.5x but doubles the price. Ad But token rates don't tell the full story without factoring in token efficiency. Opus 4.7 ended up costing 30 to 40 percent more per task than Opus 4.6 , even though both models had the same base rates. A similar pattern showed up recently with Claude Sonnet 5 . Ad DEC_D_Incontent-1 Users can trade off performance against token use through five effort settings called low, medium, high, xhigh, and max. Anthropic says Opus 5 offers better value than its predecessor at every effort level. In its prompting guide , Anthropic recommends making broad use of the "low" and "medium" settings. The company says they deliver good results with a fraction of the token use and latency while beating the same settings on earlier Opus models. For coding and agentic tasks, Anthropic still recommends starting with "xhigh." Ad Opus 5 scores slightly worse at the max effort setting than at the second-highest setting on two benchmarks, despite costing more. The drop appears on Frontier-Bench v0.1 and the Artificial Analysis Coding Agent Index. According to Anthropic's own benchmarks, Opus 5 sets records across several evaluations. On Frontier-Bench v0.1, the model hits 43.3 percent on agentic terminal coding, beating Fable 5 (33.7 percent), GPT-5.6 Sol (34.4 percent), and its predecessor Opus 4.8 (21.1 percent) by wide margins. On the knowledge work benchmark GDPval-AA v2, Opus 5 leads with an Elo score of 1,861, ahead of Fable 5 (1,747) and GPT-5.6 Sol (1,736). Ad DEC_D_Incontent-2 Opus 5 doesn't win everywhere. On agentic coding via DeepSWE v1.1, GPT-5.6 Sol leads with 72.7 percent, followed by Fable 5 (69.7 percent) and Opus 5 (68.8 percent). On health tasks and legal benchmarks, Fable 5 and Mythos 5 outperform Opus 5, respectively. Ad

中文翻译

Anthropic将Claude Opus 5定位为Fable 5的更便宜替代品,在多个基准测试中超越后者。Opus 5在代理编程和知识工作测试中领先。

核心信息

Anthropic推出Claude Opus 5,性能接近旗舰Fable 5但价格减半,在编程和知识工作基准测试中领先,并引入可调节的effort设置以优化性价比。

  • Anthropic推出Claude Opus 5,性能接近旗舰Fable 5但价格减半,在编程和知识工作基准测试中领先,并引入可调节的effort设置以优化性价比。
  • 原贴提到:Anthropic is positioning Claude Opus 5 as a cheaper alternative to Fable
  • 来源:the-decoder.com
试看内容

成为会员查看完整内容

你已经看到了这篇内容的前置整理,剩余深度部分仅对会员开放。

详细解读 信息差价值 参考来源
成为会员查看完整内容
上一篇 Midjourney V8.2 发布:专注美学提升与个性化理解 下一篇 ChatGPT Work agent现在可以登录需要账号的网站