Best AI News โ€” Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Do you think there is room for optimization? llama.cpp/qwen3.6 27b on two 6000 Blackwell

Via r/LocalLlama
Wednesday, May 20, 2026 ยท 8:48AM
Summary

Hi, i run llama.cpp inside LXC on a Proxmox server. The hardware is a recent AMD Epyc with two 6000 Blackwell MaxQ. This is my command: llama-server \ --hf-repo unsloth/Qwen3.6-27B-MTP-GGUF:BF16 --alias Qwen3.6 \ --host 0.0.0.0 --port 1337 \ --no-mmap --gpu-layers 99 \ --batch-size 6144 --ubatch-siz

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories