觉
AI觉醒星球
Awakening is here
Knowledge File / AI小生意项目库
2026-09-23 2 浏览 免费阅读

Claude Opus 5.5 以更低成本达到 Fable 5.1 的性能,并承诺少一点「Claudish」文风

Anthropic 推出新家族首个模型 Claude Opus 5.5,称其在大多数任务上达到 Claude Fable 5.1 的水平,运行成本比 Opus 5 低约 40%,token 价格下调 20%、缓存读取成本降 60%,并宣称在多个编码基准上以更低单任务成本领先 GPT-6 Astra 等竞品;同时改善写作自然度与安全护栏。

SOURCE / AI小生意项目库 MIN / 9 ACCESS / 免费阅读 POST / 2026-09-23 01:11:07

原贴

查看原文
作者:Matthias Bastian 来源站点:the-decoder.com 原贴时间:

原文

Added Artificial Analysis benchmark results Anthropic is launching Claude Opus 5.5, the first model in a new family. The company says it delivers Claude Fable 5.1-level performance while costing significantly less and running faster than its predecessor. According to Anthropic, Opus 5.5 matches Claude Fable 5.1 "on most tasks" while costing about 40 percent less to run than Opus 5. Claude Sonnet 5.5 and Haiku 5.5 are expected in the coming weeks, with similar gains in performance, efficiency, and safety. Anthropic's benchmarks show the new Opus model ahead of both Fable 5.1 and OpenAI's much more expensive GPT-6 Astra on most tasks. Anthropic says the new series primarily addresses customer feedback on cost, efficiency, and communication quality, particularly in financial services, law, and software development. Ad Anthropic prices Opus 5.5 at $4 per million input tokens and $20 per million output tokens, down from $5 and $25, respectively, for Opus 5. That's a 20 percent cut in token prices. The company also reduced cache read costs by 60 percent. Ad Anthropic says total operating costs, which account for both token prices and token usage , should be about 40 percent lower than Opus 5's. The model uses fewer tokens and generates output more than 30 percent faster. Five-hour usage limits for subscribers will increase by 20 percent. With the model's lower costs, Anthropic says those limits stretch 25 percent further overall. Users can also save a limit reset for when they need it most. Ad Anthropic is using coding benchmarks to make its case on price and performance. On FrontierCode, the company says Opus 5.5 beats OpenAI's GPT-6 Astra at about 20 percent of the cost per task. On Terminal-Bench 4.0, it claims the same performance as Astra at 40 percent of the cost. On CursorBench, it says Opus 5.5 beats GPT-5.6 Sol by 11 points at one-third of the cost. The price cut is a response to pressure from OpenAI and especially Chinese AI models, which offer lower performance but cost a fraction as much . Ad Opus 5.5 is also supposed to communicate more naturally than earlier models. Anthropic says it puts the most important information first, uses less jargon, and follows writing instructions more closely. Early testers described its writing as clearer and easier to understand, which Anthropic says makes it a better partner for long work sessions. Current Claude models have drawn plenty of criticism for their formulaic, convoluted writing, sometimes called "Claudish" . Ad Opus 5.5 is also the first Opus model with safeguards for cybersecurity, biology, and frontier LLM development that match those of Fable 5.1. When those safeguards kick in, Anthropic says requests are transparently routed to another model. Users can still find and fix bugs in their code, but most cybersecurity tasks will go to the older Opus 4.8. Requests flagged by classifiers for biology or frontier LLM development will go to Opus 5.

中文翻译

补充了 Artificial Analysis 的基准测试结果。

Anthropic 正在推出 Claude Opus 5.5,这是新家族中的首个模型。该公司称,它达到 Claude Fable 5.1 级别的性能,同时成本显著更低,运行速度也比前代更快。

据 Anthropic 称,Opus 5.5 在「大多数任务」上与 Claude Fable 5.1 相当,而运行成本比 Opus 5 低约 40%。Claude Sonnet 5.5 和 Haiku 5.5 预计将在未来几周内推出,在性能、效率和安全方面有类似提升。

Anthropic 的基准测试显示,新 Opus 模型在大多数任务上领先于 Fable 5.1 和 OpenAI 贵得多的 GPT-6 Astra。Anthropic 表示,新系列主要回应客户在成本、效率和沟通质量方面的反馈,尤其是在金融服务、法律和软件开发领域。

Anthropic 将 Opus 5.5 定价为每百万输入 token 4 美元、每百万输出 token 20 美元,低于 Opus 5 的 5 美元和 25 美元。这是 token 价格下调 20%。该公司还将缓存读取成本降低了 60%。

Anthropic 称,计入 token 价格和 token 使用量的总运营成本,应比 Opus 5 低约 40%。该模型使用的 token 更少,输出生成速度快 30% 以上。订阅用户的五小时使用限额将提高 20%。Anthropic 称,由于该模型成本更低,这些限额总体上可延伸 25%。用户还可以保存一次限额重置,留到最需要的时候使用。

Anthropic 正用编码基准测试来论证其在价格和性能上的优势。在 FrontierCode 上,该公司称 Opus 5.5 以约 20% 的单任务成本击败 OpenAI 的 GPT-6 Astra。在 Terminal-Bench 4.0 上,它声称以 40% 的成本达到与 Astra 相同的性能。在 CursorBench 上,它称 Opus 5.5 以三分之一的成本领先 GPT-5.6 Sol 11 分。此次降价是对来自 OpenAI、尤其是中国 AI 模型压力的回应,后者性能较低,但成本只是其零头。

Opus 5.5 还被期望比早期模型沟通得更自然。Anthropic 称,它把最重要的信息放在前面,使用更少的行话,并更严格地遵循写作指令。早期测试者称其写作更清晰、更易于理解,Anthropic 称这使它成为长时间工作会话中更好的伙伴。当前的 Claude 模型因公式化、绕来绕去的写作而招致不少批评,有时被称为「Claudish」。

Opus 5.5 也是首个在网络安全、生物和前沿 LLM 开发方面拥有与 Fable 5.1 相同保障措施的 Opus 模型。Anthropic 称,当这些保障措施触发时,请求会被透明地路由到另一个模型。用户仍可查找和修复代码中的 bug,但大多数网络安全任务将转到较旧的 Opus 4.8。被分类器标记为生物或前沿 LLM 开发的请求将转到 Opus 5。

核心信息

Anthropic 推出新家族首个模型 Claude Opus 5.5,称其在大多数任务上达到 Claude Fable 5.1 的水平,运行成本比 Opus 5 低约 40%,token 价格下调 20%、缓存读取成本降 60%,并宣称在多个编码基准上以更低单任务成本领先 GPT-6 Astra 等竞品;同时改善写作自然度与安全护栏。

  • Anthropic 推出新家族首个模型 Claude Opus 5.5,称其在大多数任务上达到 Claude Fable 5.1 的水平,运行成本比 Opus 5 低约 40%,token 价格下调 20%、缓存读取成本降 60%,并宣称在多个编码基准上以更低单任务成本领先 GPT-6 Astra 等竞品;同时改善写作自然度与安全护栏。
  • 原贴提到:Added Artificial Analysis benchmark results Anthropic is launching Claud
  • 来源:the-decoder.com

详细解读

这是什么信号

Anthropic 发布 Claude Opus 5.5,这是其新模型家族的首个成员,并补上了 Artificial Analysis 的第三方基准结果。官方口径是:在「大多数任务」上对齐此前更强的 Claude Fable 5.1,同时比 Opus 5 更便宜、更快。这不是一次单纯的能力升级,而是一次以成本结构为主轴的发布。

更值得注意的信号藏在降价理由里:原文明确写到,此次降价是对来自 OpenAI、尤其是中国 AI 模型压力的回应——后者性能较低,但成本只是零头。Anthropic 首次把「性价比竞争」摆到台面上,并用编码基准(FrontierCode、Terminal-Bench 4.0、CursorBench)逐个对标竞品,比的是单任务成本,而不是单纯跑分。

为什么重要

第一,价格动作是三件套而非单一折扣:token 单价输入 4 美元、输出 20 美元(较 Opus 5 的 5/25 降 20%),缓存读取成本降 60%,再叠加模型本身用更少 token、输出快 30% 以上,官方称总运营成本约低 40%。对按量付费的团队来说,真实账单降幅取决于缓存命中率和 token 用量结构,而不是价目表上的 20%。

第二,订阅侧同步扩容:五小时限额提高 20%,因成本下降整体可延伸 25%,并允许保存一次限额重置。这是把模型效率提升直接转化为用户可用额度,而不是只让 API 客户受益。

第三,写作质量被当成一级卖点。原文提到当前 Claude 模型因公式化、绕来绕去的「Claudish」文风饱受批评,Opus 5.5 强调信息前置、少行话、更遵循写作指令。这说明在长会话与文档类场景里,瓶颈已从推理能力转向表达质量。

第四,安全护栏出现分层路由:网络安全、生物、前沿 LLM 开发三类防护对齐 Fable 5.1,触发后请求被透明地转到其他模型——多数网络安全任务去 Opus 4.8,生物与前沿 LLM 开发类请求去 Opus 5。这是能力「按敏感度分渠道供给」的一次公开化操作。

对谁有价值

最直接受益的是高频调用 API 的开发者与初创团队:缓存读取降 60% 对大量复用系统提示词、长上下文检索的应用影响最大。其次是订阅重度用户,限额提升与可保存的重置权对长时间编码会话有实际意义。再次是金融、法律、软件开发等被点名行业——这些场景既贵又要求文字表达清晰,正是本次成本与文风双改动的目标人群。企业采购方也能借此重新谈判模型预算。

可以怎么行动

  • 用自己的任务集复测,而不是照搬厂商基准:原文列的是 FrontierCode、Terminal-Bench 4.0、CursorBench,对照口径是「同等性能下的单任务成本」,建议按此搭建内部选型表。
  • 把评估指标从「每百万 token 价格」换成「每任务成本」,把缓存命中率、平均 token 用量一起纳入模型,否则算不出那 40%。
  • 针对安全分流提前设计路由与回退:网络安全类任务会落到 Opus 4.8,生物与前沿 LLM 开发类请求会转到 Opus 5,工作流需要能承受模型切换带来的行为差异。
  • 把写作类 prompt 做 A/B,验证「重要信息前置、少行话、更遵循指令」在实际业务文本上是否成立。
  • 关注 Sonnet 5.5 与 Haiku 5.5,官方称未来几周内推出,并会在性能、效率、安全上有类似提升——分层产品线决定最终部署成本组合。

风险或限制

首先,性能对标主要来自厂商自述,且原文措辞是「在大多数任务上」匹配 Fable 5.1,留有余地,不同任务分布下结论可能不同。其次,成本下降是价目表、缓存折扣与 token 用量三者叠加的估算,实际节省强依赖具体工作负载。第三,安全分流意味着同一模型名下的能力并不一致,被路由走的任务既可能换模型也可能换表现,需要单独评估。第四,降价动机是对 OpenAI 与中国模型的价格压力作出的反应,说明这一价格位置具有竞争性而非结构性,后续存在再次变动的可能。

信息差价值

这条内容的真正价值,不只是“有人发布了一个新功能”,而是它揭示了 the-decoder.com 背后的产品方向、工作流变化或竞争信号。对 OPC 来说,这种信息可以转化成持续追踪的栏目选题。

如果把《Claude Opus 5.5 以更低成本达到 Fable 5.1 的性能,并承诺少一点「Claudish」文风》放到你的内容系统里,它最大的价值在于帮助读者更快看懂“为什么值得关注”,而不是只看到一条碎片化动态。

参考来源

AI SUMMARY

这篇文章回答了什么

Claude Opus 5.5 以更低成本达到 Fable 5.1 的性能,并承诺少一点「Claudish」文风主要讲什么?

Anthropic 推出新家族首个模型 Claude Opus 5.5,称其在大多数任务上达到 Claude Fable 5.1 的水平,运行成本比 Opus 5 低约 40%,token 价格下调 20%、缓存读取成本降 60%,并宣称在多个编码基准上以更低单任务成本领先 GPT-6 Astra 等竞品;同时改善写作自然度与安全护栏。

这篇文章最值得关注的要点是什么?

Anthropic 推出新家族首个模型 Claude Opus 5.5,称其在大多数任务上达到 Claude Fable 5.1 的水平,运行成本比 Opus 5 低约 40%,token 价格下调 20%、缓存读取成本降 60%,并宣称…;原贴提到:Added Artificial Analysis benchmark results Anthropic is launching Claud;来源:the-decoder.com

这篇文章和哪些AI专题相关?

它适合放在AI副业、AI工具、AI超级个体专题里阅读。 关联原因:这篇内容命中「项目、小生意、变现」等主题信号。;这篇内容命中「模型、Claude」等主题信号。;这篇内容命中「效率」等主题信号。

阅读这篇文章建议先理解哪些关键词?

建议先理解AI工具、工具、自动化、模型、Cursor这些关键词,再结合正文判断工具、机会或风险是否值得进入自己的工作流。

上一篇 Sam Altman 称 GPT-6 Sol 和 Luna 按任务定价在市场上没有对手 下一篇 Claude Opus 5.5 现已在 GitHub Copilot 中可用