Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

MI50s Qwen 3.6 27B @52.8 tps TG @1569 tps PP (no MTP, no Quant)

Via r/LocalLlama
Wednesday, May 13, 2026 · 7:08PM
Summary

TL;DR Results from the title are for single inference with 2 prompt of 1k and 15k tokens. So no MTP (as it’s slower for big prompt), no DFlash (working too but slower for big prompt), no quant used (full precision wanted) and the results are pretty good for a 2018 card. (Bench has been done with TP8

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories