Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

ubergarm/Kimi-K2.6-GGUF Q4_X now available

Via r/LocalLlama
Monday, Apr 20, 2026 · 8:38PM
Summary

Big thanks to jukofyork and AesSedai today giving me some tips to patch and quantize the "full size" Kimi-K2.6 "Q4_X". It runs on both ik and mainline llama.cpp if you have over ~584GB RAM+VRAM... I'll follow up with imatrix for anyone else making custom quants, and some smaller quants that run on i

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories