https://huggingface.co/unsloth/gemma-4-E2B-it-GGUF https://huggingface.co/unsloth/gemma-4-26B-A4B-it-GGUF by u/danielhanchen: We just updated them again in response to: kv-cache : support attention rotation for heterogeneous iSWA https://github.com/ggml-org/llama.cpp/pull/21513 CUDA: check for buffe