I'm struggling to find the right setup with llama.cpp. Ubuntu 26.04, AMD 5900X, 128GB DDR4-3600, R9700s are running PCIe x8/Gen4 Model config: [Qwen3.6-27B] mmproj = /models/Qwen3.6-27B-mmproj-BF16.gguf model = /models/mtp/Qwen3.6-27B-Q8_0.gguf alias = Qwen3.6-27B ctx-size = 180000 threads = 12 temp