Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Qwen3.5 122B-A10B · ROCmFP4 iMatrix

Via r/LocalLlama
Thursday, Jul 16, 2026 · 2:45AM
Summary

Hola Strix and AMD stacker frendios. Read the Lineage and Credits, this uses charlie12345/ROCmFPX, won't work on native llama.cpp yet. 122B total · 10B active · 60.70 GiB · 28.50 tok/s MTP-off · BF16 KLD 0.041366 · Decode 28.505 Decode speed + 36.89% faster Size - 13.47gb smaller

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories