Google’s TurboQuant is getting all the attention for KV cache compression (6× smaller, zero loss). Cool. But the weights are still eating your VRAM. TurboQuant-v3 fixes that: • Group-wise INT4 + AWQ scaling + protected FP16 outliers + optional SVD correction • ~4× memory reduction • 2–3× speedup via