Pure C implementation of the TurboQuant paper (ICLR 2026) for KV cache compression in LLM inference. Key vectors compressed to 1 bit via randomized Hadamard transform + sign hashing. Attention via XOR + popcount. Values independently quantized to Q4 or Q2. Total K+V: 4.9x–7.1x compression on Gemma 3