Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

The exact KV cache usage of DeepSeek V4

Via r/LocalLlama
Sunday, Apr 26, 2026 · 6:19AM
Summary

Figure 1 of DSV4 paper seems to imply that DSV3.2 uses ~50GB at 1m context and DSV4 uses ~5GB: https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro/blob/main/DeepSeek_V4.pdf From my own calculations, the correct FP16 KV cache at 1m context should be: Model Params 128k 160k 1m KV% V3.x 671B 8.58GiB 10.

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories