Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Get faster qwen 3.6 27b

Via r/LocalLlama
Wednesday, May 6, 2026 · 11:33PM
Summary

Using 100k context with 3090 with MTP GGUF and getting 50 t/s on llama.cpp Thought I would knowledge share Use https://huggingface.co/RDson/Qwen3.6-27B-MTP-Q4_K_M-GGUF And am17an commit /media/adam/D_DRIVE/LLM/llama-cpp-am17an/build/bin/llama-server -m "/media/Qwen3.6-27B-Q4/Qwen3.6-27B-MTP-Q4_K_M.g

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories