觉
AI觉醒星球
Awakening is here
Knowledge File / AI小生意项目库
2026-10-08 1 浏览 免费阅读

Claude Haiku 5.5 发布并大幅降价,证明 AI 价格战远未结束

Anthropic 发布 Claude Haiku 5.5,性能较前代大幅提升,多数请求 token 价格最高降 90%,并同步下调 Sonnet 5.5 缓存读取成本。该模型已在 AWS、Google Cloud 和 Azure 上线,但新 tokenizer 可能削弱实际节省幅度。

SOURCE / AI小生意项目库 MIN / 9 ACCESS / 免费阅读 POST / 2026-10-08 02:49:23

原贴

查看原文
作者:Matthias Bastian 来源站点:the-decoder.com 原贴时间:

原文

Anthropic has released Claude Haiku 5.5, its fastest and most affordable small model, built for high-volume tasks like data queries and customer support. Haiku 5.5 shows major benchmark gains over Haiku 4.5 at up to 90 percent lower token prices. It also beats OpenAI's budget model GPT-6 Luna in areas like agentic coding. Haiku 5.5 is available now on AWS, Google Cloud, and Azure. Anthropic is also cutting Sonnet 5.5 cache read costs in half and rolling out monthly API credits for subscribers. Anthropic has released Claude Haiku 5.5, the company's fastest and most affordable small model to date. Benchmark results show a major performance jump over its predecessor, and Anthropic is also cutting prices for Sonnet 5.5. Haiku 5.5 is designed for high-volume, cost-sensitive tasks like summarization, database queries, classification, and live customer support, according to Anthropic. On average, the model costs about 75 percent less than Haiku 4.5. For requests with prompts up to 100,000 tokens, which Anthropic says account for roughly 90 percent of all previous Haiku requests, prices drop by up to 90 percent. Prompts longer than 100,000 tokens cost five times as much. Anthropic points out that Haiku 5.5 uses an updated tokenizer that consumes slightly more tokens per task than its predecessor. The same thing happened with the Opus 4.x models , where token usage jumped about 30 percent from the tokenizer change alone. Real-world savings are likely smaller than the per-token prices suggest. Ad Haiku 5.5 scores 1,620 on the knowledge benchmark GDPval-AA v2.1, more than double the 735 its predecessor managed. On Humanity's Last Exam, it hits 45.9 percent without tools and 57.4 percent with tools, up from 10.2 and 18.7 percent. Ad The biggest jump is in computer use , where the model operates a computer on its own. Since computer use burns through large amounts of tokens, a cheap, fast model like Haiku 5.5 is a natural fit. It scores 72.4 percent on OSWorld-2.1, up from 15.7 percent. On the agentic coding benchmark Terminal-Bench 4.0, it reaches 39.2 percent while Haiku 4.5 scored zero. Anthropic also lists OpenAI's budget model GPT-6 Luna as a comparison, and Haiku 5.5 leads across every tested category. Sonnet 5.5 reference scores show, however, that Haiku 5.5 still falls well behind Anthropic's larger model. Ad Haiku 5.5 is the first Haiku-class model with adjustable reasoning levels, letting users balance cost against quality. Anthropic says the model works best for narrowly scoped tasks like compaction, summarization, or sub-agent work. For complex agentic coding, Sonnet 5.5 and Opus 5.5 remain the better picks. Cybersecurity safeguards are tighter than on the predecessor but allow a broader range of defensive tasks than Sonnet 5.5, partly because the model is less capable overall. Penetration testing stays blocked. Organizations with broader needs can apply for Anthropic's verification programs for life sciences and cybersecurity . Ad Haiku 5.5 is available now across all platforms, including Amazon Web Services, Google Cloud, and Microsoft Azure. Ad

中文翻译

Anthropic 发布了 Claude Haiku 5.5,这是该公司迄今为止最快、最实惠的小型模型,专为数据查询和客户支持等高吞吐任务打造。Haiku 5.5 在基准测试上较 Haiku 4.5 有重大提升,token 价格最高降低 90%。它在智能体编程等领域也击败了 OpenAI 的廉价模型 GPT-6 Luna。Haiku 5.5 现已在 AWS、Google Cloud 和 Azure 上线。Anthropic 还将 Sonnet 5.5 的缓存读取成本减半,并为订阅用户推出月度 API 额度。

Anthropic 发布了 Claude Haiku 5.5,这是该公司迄今为止最快、最实惠的小型模型。基准测试结果显示其性能较前代大幅跃升,Anthropic 同时也在下调 Sonnet 5.5 的价格。Haiku 5.5 面向高吞吐、成本敏感的任务,如摘要、数据库查询、分类和实时客户支持。平均而言,该模型成本比 Haiku 4.5 低约 75%。对于提示最多 100,000 token 的请求(Anthropic 称约占此前所有 Haiku 请求的 90%),价格最多下降 90%。超过 100,000 token 的提示成本是五倍。Anthropic 指出,Haiku 5.5 使用更新的 tokenizer,每个任务消耗的 token 比前代略多。Opus 4.x 模型也发生过同样情况,仅 tokenizer 变化就使 token 使用量增加约 30%。现实世界节省可能小于每 token 价格所暗示的。Haiku 5.5 在知识基准 GDPval-AA v2.1 上得分 1,620,是前代 735 的两倍多。在 Humanity's Last Exam 上,无工具时达到 45.9%,有工具时达到 57.4%,此前分别为 10.2% 和 18.7%。最大跃升在计算机使用方面,模型可自行操作计算机。由于计算机使用消耗大量 token,像 Haiku 5.5 这样便宜快速的模型天然适合。它在 OSWorld-2.1 上得分 72.4%,此前为 15.7%。在智能体编程基准 Terminal-Bench 4.0 上,它达到 39.2%,而 Haiku 4.5 得分为零。Anthropic 还列出 OpenAI 的廉价模型 GPT-6 Luna 作比较,Haiku 5.5 在所有测试类别中领先。然而 Sonnet 5.5 参考分数显示,Haiku 5.5 仍远远落后于 Anthropic 的更大模型。Haiku 5.5 是首个具有可调节推理级别的 Haiku 级模型,让用户平衡成本与质量。Anthropic 称该模型最适合范围狭窄的任务,如压缩、摘要或子智能体工作。对于复杂的智能体编程,Sonnet 5.5 和 Opus 5.5 仍是更好选择。网络安全保障比前代更严格,但比 Sonnet 5.5 允许更广泛的防御性任务,部分因为该模型整体能力较弱。渗透测试仍被阻止。有更广泛需求的组织可申请 Anthropic 的生命科学和网络安全验证计划。Haiku 5.5 现已在所有平台上线,包括 Amazon Web Services、Google Cloud 和 Microsoft Azure。

核心信息

Anthropic 发布 Claude Haiku 5.5,性能较前代大幅提升,多数请求 token 价格最高降 90%,并同步下调 Sonnet 5.5 缓存读取成本。该模型已在 AWS、Google Cloud 和 Azure 上线,但新 tokenizer 可能削弱实际节省幅度。

  • Anthropic 发布 Claude Haiku 5.5,性能较前代大幅提升,多数请求 token 价格最高降 90%,并同步下调 Sonnet 5.5 缓存读取成本。该模型已在 AWS、Google Cloud 和 Azure 上线,但新 tokenizer 可能削弱实际节省幅度。
  • 原贴提到:Anthropic has released Claude Haiku 5.5, its fastest and most affordable
  • 来源:the-decoder.com

详细解读

Anthropic 发布 Claude Haiku 5.5,并同步下调 Sonnet 5.5 缓存读取成本、为订阅用户推出月度 API 额度。表面看是一次常规小模型迭代,实际是 AI 价格战继续升级的明确信号:小模型不只是“便宜版本”,而是被推到高吞吐、智能体计算机使用等真实工作负载的前线。

为什么重要:Haiku 5.5 在 GDPval-AA v2.1 上得分 1,620(前代 735),Humanity's Last Exam 有工具达到 57.4%,OSWorld-2.1 从 15.7% 提升到 72.4%,Terminal-Bench 4.0 从 0 到 39.2%。同时,多数请求价格最多降 90%。这意味着过去因成本被挡在门外的批量任务,如摘要、分类、数据库查询、实时客服,以及消耗大量 token 的计算机使用,现在可以重新评估是否用模型跑。

对谁有价值:一是需要大规模调用模型的产品和运营团队,单位成本下降直接改变毛利和定价空间;二是做 AI 智能体、浏览器自动化、桌面操作的产品,Haiku 5.5 的速度与价格更适合子智能体、压缩、摘要等窄任务;三是关注云平台选型的团队,因为它已在 AWS、Google Cloud、Azure 上线。

可以怎么行动:先用 Haiku 5.5 替换原先跑在 Haiku 4.5 或更贵模型上的高吞吐任务,做 A/B 对比;对长提示特别小心,超过 100,000 token 的请求价格是普通档的五倍,适合把任务拆成短提示;利用可调推理级别,对窄任务设低成本档,对质量敏感任务设高推理档;复杂智能体编程仍应路由到 Sonnet 5.5 或 Opus 5.5。

风险与限制:Anthropic 自己指出,新 tokenizer 每个任务消耗的 token 略多,Opus 4.x 曾因 tokenizer 变化导致 token 使用量增加约 30%,所以实际节省可能小于每 token 价格所示。Haiku 5.5 整体能力仍低于 Sonnet 5.5,且渗透测试被阻止,网络安全防御任务范围虽然比 Sonnet 5.5 更宽,但也是因为模型能力较弱。把关键决策或高风险自动化完全交给小模型,仍需人工审核与安全边界。

信息差价值

这条内容的真正价值,不只是“有人发布了一个新功能”,而是它揭示了 the-decoder.com 背后的产品方向、工作流变化或竞争信号。对 OPC 来说,这种信息可以转化成持续追踪的栏目选题。

如果把《Claude Haiku 5.5 发布并大幅降价,证明 AI 价格战远未结束》放到你的内容系统里,它最大的价值在于帮助读者更快看懂“为什么值得关注”,而不是只看到一条碎片化动态。

参考来源

AI SUMMARY

这篇文章回答了什么

Claude Haiku 5.5 发布并大幅降价,证明 AI 价格战远未结束主要讲什么?

Anthropic 发布 Claude Haiku 5.5,性能较前代大幅提升,多数请求 token 价格最高降 90%,并同步下调 Sonnet 5.5 缓存读取成本。该模型已在 AWS、Google Cloud 和 Azure 上线,但新 tokenizer 可能削弱实际节省幅度。

这篇文章最值得关注的要点是什么?

Anthropic 发布 Claude Haiku 5.5,性能较前代大幅提升,多数请求 token 价格最高降 90%,并同步下调 Sonnet 5.5 缓存读取成本。该模型已在 AWS、Google Cloud 和 Azure 上线…;原贴提到:Anthropic has released Claude Haiku 5.5, its fastest and most affordable;来源:the-decoder.com

这篇文章和哪些AI专题相关?

它适合放在AI副业、AI工具、Agent工作流专题里阅读。 关联原因:这篇内容命中「项目、小生意、变现」等主题信号。;这篇内容命中「模型、Claude」等主题信号。;这篇内容命中「智能体」等主题信号。

阅读这篇文章建议先理解哪些关键词?

建议先理解AI工具、工具、自动化、模型、Cursor这些关键词,再结合正文判断工具、机会或风险是否值得进入自己的工作流。

上一篇 Anthropic 发布 Claude Haiku 5.5,Artificial Analysis 智能指数得分 43 下一篇 NVIDIA 与 Microsoft 推出 RTX Spark 平台并宣布 MXC 让 AI Agent 落地 Windows PC
北竹游乐场 免费玩小游戏 免费玩