Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

HIP: use hipBLAS for dense prefill on gfx900, keep MMQ for MoE by DEV-DUFORD · Pull Request #24588 · ggml-org/llama.cpp

Via r/LocalLlama
Tuesday, Jun 30, 2026 · 5:27PM
Summary

Overall Performance Gains: Qwen3.5 4B: +36.1% Qwen3.6 27B: +18.9% Gemma4 12B: +65.1% Overall average: ~40% Only for gfx900 related GPUs: Vega GPU, codename vega10, including Radeon Vega Frontier Edition, Radeon RX Vega 56/64, Radeon RX Vega 64 Liquid, Radeon Pro Vega 48/56/64/64X, Radeon Pro WX 8200

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories