TL;DR; GLM-5.2 Q1_S beats Qwen 3.6 27B Q8, both run at KV Q8 edit: GLM run a K & V Q8, Qwen run with KV cache at full FP16., with preserve thinking on. Disclaimer: This is a hobby/amateur comparison with n=1, so go easy on it. I just thought it would be fun to share. The Context and The Task Some ti