Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

DeepSeek V4 Flash, up to 32 tok/s on AMD Ryzen AI MAX+ 395

Via r/LocalLlama
Tuesday, Jul 28, 2026 · 3:00PM
Summary

Hey fellow llamas. we have something new for Strix Halo owners we thought would be useful to share. i'll keep it short: We were able to fit DeepSeek V4 Flash plus its speculative draft on a single Ryzen AI MAX+ 395 with 128 GB of unified memory, and got it to a usable decode rate. Blog post with all

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories