https://github.com/ggml-org/llama.cpp/pull/23966 https://github.com/ggml-org/llama.cpp/pull/22716 Use llama.cpp version.