Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

MTP hyperparameter search

Via r/LocalLlama
Thursday, Jun 11, 2026 · 3:37AM
Summary

TLDR; I only got a 6% improvement on tokens/sec over naïve parameters. I was messing around and ran a hyperparameter search with optuna over the MTP and speculative decoding options of llama-server for Qwen3.6 27b on strix halo. Here's the very rough python script (created by Qwen): https://gist.git

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories