AI觉醒星球
Awakening is here
Knowledge File / AI小生意项目库
2026-08-01 0 浏览 会员

新Deepseek Flash模型以约60%更低成本媲美OpenAI的GPT-5.6 Luna

Deepseek发布V4 Flash "0731",性能得分50,仅比OpenAI的GPT-5.6 Luna低1分,但任务成本低约60%。凭借98%的缓存折扣、减少12%的token消耗,模型在代理任务和GDPval等基准上大幅进步,且幻觉更少。架构不变,权重以MIT许可在Hugging Face开放。

SOURCE / AI小生意项目库 MIN / 9 ACCESS / 会员 POST / 2026-08-01 00:39:25

原贴

查看原文
作者:Thomas Joos 来源站点:the-decoder.com 原贴时间:

原文

Deepseek has released V4 Flash "0731," a major upgrade to its budget AI model. According to the Artificial Analysis Intelligence Index , the new version scores 50 points, ten more than the previous V4 Flash that launched in April 2026. That puts it just one point behind OpenAI's budget model GPT-5.6 Luna, but it costs about 60 percent less per task, even after OpenAI's 80 percent price cut . A big reason for the gap is Deepseek's 98 percent cache discount , well above the industry-standard 90 percent. The model also uses 12 percent fewer tokens than its predecessor. The model improves across every tested category compared to the previous version, with the biggest gains in agentic tasks. On GDPval, a benchmark designed to test models on complex real-world office work , it climbs from 1,189 to 1,559 Elo points. It also hallucinates less often. The architecture stays the same: 284 billion total parameters, 13 billion active, with a one-million-token context window. The model weights are available under an MIT license on Hugging Face . Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section.

中文翻译

Deepseek发布了V4 Flash "0731",这是其经济型AI模型的重大升级。根据Artificial Analysis Intelligence Index,新版本得分50分,比2026年4月发布的上一代V4 Flash高出10分。这使得它仅落后OpenAI经济型模型GPT-5.6 Luna 1分,但每个任务的成本要低约60%,即使OpenAI已降价80%。造成这一差距的一大原因是Deepseek的98%缓存折扣,远高于行业标准的90%。该模型使用的token也比前代少12%。与上一版本相比,该模型在每一个测试类别上都有提升,其中在代理任务上进步最大。在GDPval(一个旨在测试模型在复杂真实世界办公工作中表现的基准)上,它从1,189 Elo分上升到1,559 Elo分。它的幻觉也减少了。架构保持不变:总参数2840亿,激活参数130亿,上下文窗口为100万token。模型权重以MIT许可在Hugging Face上提供。订阅THE DECODER可获得无广告阅读、每周AI通讯、每年六次的独家"AI Radar"前沿报告、完整档案访问权限以及评论区的访问权限。

核心信息

Deepseek发布V4 Flash "0731",性能得分50,仅比OpenAI的GPT-5.6 Luna低1分,但任务成本低约60%。凭借98%的缓存折扣、减少12%的token消耗,模型在代理任务和GDPval等基准上大幅进步,且幻觉更少。架构不变,权重以MIT许可在Hugging Face开放。

  • Deepseek发布V4 Flash "0731",性能得分50,仅比OpenAI的GPT-5.6 Luna低1分,但任务成本低约60%。凭借98%的缓存折扣、减少12%的token消耗,模型在代理任务和GDPval等基准上大幅进步,且幻觉更少。架构不变,权重以MIT许可在Hugging Face开放。
  • 原贴提到:Deepseek has released V4 Flash "0731," a major upgrade to its budget AI
  • 来源:the-decoder.com
试看内容

成为会员查看完整内容

你已经看到了这篇内容的前置整理,剩余深度部分仅对会员开放。

详细解读 信息差价值 参考来源
成为会员查看完整内容
上一篇 Thinking Machines押注效率而非规模,推出第二款模型Inkling Small 下一篇 欧盟为AI超级工厂筹集300亿欧元,而美国科技巨头轻松花费20倍以上