*general not whatever that word is. Assume I'm talking about Qwen3.6b-27b for coding. I hear a lot about quantisizing the model but almost no opinions on the KV cache for this model. EDIT: Btw thanks everyone, I'm in awe of how much I learn from this sub every day.