Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Minimum VRAM GPU to run DeepSeek-V4-Flash-0731 Q4_K_XL at around 30 t/s ?

Via r/LocalLlama
Friday, Jul 31, 2026 · 7:37PM
Summary

Hello guys, I'm curious about running DeepSeek-V4-Flash-0731 locally. Since it’s a Mixture of Experts (MoE) model with only 13B active parameters, I was hoping the VRAM requirements might be manageable. Did someone tried out in some reasonable GPU sizes up to 48GB VRAM? Thanks for the feedback!

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories