Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

PSA: llama.cpp now loads MTP tensors by default for any draft-mtp arch, even with MTP disabled

Via r/LocalLlama
Wednesday, Jul 29, 2026 · 6:45PM
Summary

If your GGUF has MTP/NextN tensors baked in (GLM-5.2, hy_v3, qwen35moe, step35, etc.), recent llama.cpp builds load them by default — even if you never pass --spec-type draft-mtp. Before, they were skipped unless you actually enabled speculative decoding. Most community GGUFs bundle the MTP block by

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories