Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

DeepSeek-V4-Flash in MXFP4 is too slow on CPU

Via r/LocalLlama
Sunday, Jul 5, 2026 · 7:35AM
Summary

I have an old Xeon rig with 512Gb of 4-channel DDR4 2133 memory and E5-2699v4 processor. For GPU I have GTX 1060 with 6Gb of VRAM, so I use CPU only mode. I can run GLM 5.2 with 40B active parameters in Q4_K_XL at 1.8 t/s, but as you can understand it is too slow. So I wanted to give a try to a new

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories