Best AI News โ€” Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Benchmarking the new b9200 update: Optimizing Qwen 3.6 27B mtp for Hermes Agent on a single RTX 3090

Via r/LocalLlama
Monday, May 18, 2026 ยท 12:20AM
Summary

UPDATED (POST b9200) Okay, here is the updated version using the new Qwen 3.6 27B mtp gguf from Unsloth, running it as the backend for the hermes agent. While dialing it in, I noticed that the currently recommended Unsloth mtp flags actually bottleneck performance and tank draft acceptance rates for

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories