跳到正文
RCreddit.com·

DeepSeek is ruthless

AI 摘要

DeepSeek has introduced DeepSeek-V4.1-Flash, a new method significantly compressing the memory requirements for the KV-value cache. This innovation from DeepSeek and other Chinese labs is seen as aggressively reducing inference costs. Such advancements could pose a challenge for companies like OpenAI and Anthropic, potentially making it difficult for them to recoup the substantial investments made in developing their leading models.

为什么是这条

This report highlights DeepSeek's DeepSeek-V4.1-Flash as a specific example of Chinese labs aggressively reducing inference costs, unlike the broader cost-reduction efforts from Western counterparts.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

发布当时偏移:UTC+02026年9月13日 14:41 UTC

收录当时偏移:UTC+02026年9月13日 15:00 UTC

发布
2026年9月13日 14:41
收录
2026年9月13日 15:00
来源类型
开发者社区
档位
社区
信源状态
正常

档位是按信源手工设定的编辑判断,不是逐条打分。

DeepSeek has published with DeepSeek-V4.1-Flash a new method that compresses the memory need for the KV-value cache very much.

I pondered about the implications of this and they are not very good for OpenAI and Anthropic.

This means that the models can have much larger contexts and serving requests will be much less memory intensive. As a result, inference gets cheaper.

Inference getting cheaper, requiring less memory and with better models means that the advantage OpenAI and Anthropic has in securing compute gets less meaningful.

It seems to me that DeepSeek and other Chinese labs are ruthlessly pushing down the cost of inference, which will make it difficult to impossible for OpenAI and Anthropic to recover all the money spent of creating their top models.

来源·reddit.com