Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

LoRA over GGUF: Train DeepSeek-V4-Flash in 90G VRAM

Via r/LocalLlama
Tuesday, Jul 28, 2026 · 4:49PM
Summary

https://github.com/woct0rdho/transformers5-qwen3.5-recipe An update on my progress with low-VRAM LoRA training over GGUF base model: Now we can train DeepSeek-V4-Flash (284B-A13B) in 90 GiB VRAM, with no CPU offloading. On Strix Halo it runs at 19 s/it. All the WTF parts - sliding attention, CSA, HC

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories