DeepSeek V4: the most expected open-source model ever released, and the quietest landing
After 15 months of incremental updates, leaks, and rumored leaks, DeepSeek released version 4. It arrived without the fanfare R1 and R1-preview commanded in early 2025. That quiet reception is the most interesting thing about the release. A few months ago, the same model would have dominated the...
After 15 months of incremental updates, leaks, and rumored leaks, DeepSeek released version 4. It arrived without the fanfare R1 and R1-preview commanded in early 2025. That quiet reception is the most interesting thing about the release. A few months ago, the same model would have dominated the cycle. Now the headlines include a mixture of open and closed models trained on multiple infrastructure providers. The infrastructure conversation has crowded out the model conversation, and DeepSeek v4 is the first major open release to land in that climate. For teams running these models in production, that's the right context. The architecture changes in v4 are engineering wins more than capability leaps, and engineering wins is what matters when you're paying for serving. Beyond the model itself, NVIDIA and Lambda co-design infrastructure and optimize performance to further reduce the cost per token on open models like DeepSeek V4 as proven by the latest MLPerf Inference V6 results.Source: Lambda Labs — Published — Category: Models