Best AI News โ€” Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Used over a million tokens in three separate sessions to test Qwen 3.6 35b (new Multi-token Prediction version)

Via r/LocalLlama
Friday, May 15, 2026 ยท 6:20AM
Summary

In my opinion, MTP models are 100% game changer for local LLMs. In terms of speed, I was getting around 1.5x the tok/sec of previous tests. The project was a test - building a full iterative step-by-step pygame; a small mystery dungeon-style game. At first I set 100-200k context and raised it to 300

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories