New Deepseek Flash model matches OpenAI's GPT-5.6 Luna at roughly 60 percent lower cost
The Decoder··作者 Thomas Joos
资讯摘要
New Deepseek Flash model matches OpenAI's GPT-5.6 Luna at roughly 60 percent lower cost Deepseek has released V4 Flash "0731," a major upgrade to its budget AI model. According to the Artificial Analysis Intelligence Index, the new version scores 50 points, ten more than the previous V4 Flash that launched in April 2026. That puts it just one point behind OpenAI's budget model GPT-5.6 Luna, but it costs about 60 percent less per task, even after OpenAI's 80 percent price cut. A big reason for the gap is Deepseek's 98 percent cache discount, well above the industry-standard 90 percent.

资讯正文
新的 Deepseek Flash 模型以大约低 60% 的成本匹配 OpenAI 的 GPT-5.6 Luna
Deepseek 发布了 V4 Flash“0731”,这是其低成本 AI 模型的一次重大升级。根据 Artificial Analysis Intelligence Index,新版本得分为 50 分,比 2026 年 4 月推出的上一代 V4 Flash 高出 10 分。这意味着它只比 OpenAI 的低成本模型 GPT-5.6 Luna 低 1 分,但即便是在 OpenAI 已将价格下调 80% 之后,它每项任务的成本仍然低了大约 60%。造成这一差距的一个重要原因是 Deepseek 的 98% 缓存折扣,远高于行业标准的 90%。与前代相比,该模型使用的 token 也减少了 12%。
与上一版本相比,该模型在所有测试类别中都有提升,其中在 agentic 任务上的进步最大。在 GDPval 上,这一基准用于测试模型在复杂真实办公工作中的表现,它的 Elo 分数从 1,189 提升到 1,559。它的幻觉也更少了。架构保持不变:总参数 2840 亿,激活参数 130 亿,拥有 100 万 token 的上下文窗口。模型权重已在 Hugging Face 上以 MIT 许可证公开。
来源与参考
收录于 2026-08-01