Best AI News โ€” Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Luce DFlash + PFlash on AMD Strix Halo: Qwen3.6-27B at 2.23x decode and 3.05x prefill vs llama.cpp HIP

Via r/LocalLlama
Tuesday, May 12, 2026 ยท 6:09PM
Summary

Hey fellow Llamas, keeping it short. We just shipped DFlash and PFlash support for the AMD Ryzen AI MAX+ 395 iGPU (gfx1151, Strix Halo, 128 GiB unified memory). Same Luce DFlash stack from the RTX 3090 post a couple weeks back, now running on the consumer AMD APU class. Repo: https://github.com/Luce

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories