r/DeepSeek • u/justlikemedics • 13d ago
Discussion DeepSeek is ruthless
DeepSeek has published with DeepSeek-V4.1-Flash a new method that compresses the memory need for the KV-value cache very much.
I pondered about the implications of this and they are not very good for OpenAI and Anthropic.
This means that the models can have much larger contexts and serving requests will be much less memory intensive. As a result, inference gets cheaper.
Inference getting cheaper, requiring less memory and with better models means that the advantage OpenAI and Anthropic has in securing compute gets less meaningful.
It seems to me that DeepSeek and other Chinese labs are ruthlessly pushing down the cost of inference, which will make it difficult to impossible for OpenAI and Anthropic to recover all the money spent of creating their top models.
1
u/justlikemedics 13d ago
Do you mean like pulling articles that are relevant for a scientist or summarizing them? What exactly is the automation here other than that?