Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Qwen3.5-4B GGUF quants comparison (KLD vs speed) - Lunar Lake

Via r/LocalLlama
Monday, Apr 6, 2026 · 6:04AM
Summary

I wanted to know which type of quant is the best on this laptop (Intel 258V - iGPU 140V 18GB), so I tested all these small quants hoping that it generalizes to bigger models: Winners in bold (KLD≤0.01) Uploader Quant tk/s KLD GB KLD/GB* mradermacher* Q4_0 28.97 0.052659918 2.37 0.04593 mradermacher_

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories