Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Giving GLM-5.2 a spin locally on CPU only! (poor man's rig for big models)

Via r/LocalLlama
Thursday, Jun 18, 2026 · 9:40PM
Summary

This is the UD-Q2-K_XL quant. Hardware is: Model: Dell PowerEdge R740 CPU: Dual Xeon 6248R (24 cores each) RAM: 768 GB (All memory channels populated) I'm using ik_llama.cpp which provides some significant performance improvements over the base llama.cpp for CPU-only inference. Unfortunately, we dua

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories