Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

ggml-cpu: Optimized x86 and generic cpu q1_0 dot (follow up) by pl752 · Pull Request #21636 · ggml-org/llama.cpp

Via r/LocalLlama
Tuesday, Apr 21, 2026 · 11:41AM
Summary

Available b8858 onwards. This is optimized CPU version so faster t/s now. (Just tested on my old weak laptop(16GB DDR3 RAM). Before : 0.3 t/s & After : 1.7 t/s. Obviously I didn't get expected boost as my laptop don't have AVX or AVX512 support. I'll be checking on my new laptop this week.) FYI Meta

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories