Best AI News โ€” Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Autoresearch on Qwen3.5-397B, 36 experiments to reach 20.34 tok/s on M5 Max, honest results

Via r/LocalLlama
Monday, Mar 30, 2026 ยท 4:01AM
Summary

I spent the past week trying to push Qwen3.5-397B faster on my M5 Max 128GB. Dan Woods' (@danveloper) original baseline was 4.36 tok/s on M3 Max. On M5 Max the starting point was already 10.61 tok/s due to better hardware. My optimizations pushed it to 20.34 tok/s, roughly 2x through software alone,

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories