Best AI News โ€” Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Computer build using Intel Optane Persistent Memory - Can run 1 trillion parameter model at over 4 tokens/sec

Via r/LocalLlama
Monday, May 11, 2026 ยท 7:54PM
Summary

As the title states, my build is indeed able to run a 1 trillion parameter model (in this case Kimi K2.5) locally at ~4 tokens/second. I thought r/LocalLLaMA would be interested in the build due to that stat line, and also due to the inclusion of an unusual part, Intel Optane Persistent Memory, whic

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories