Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Cheapest setup for >10 tok/sec for 120B dense LLM

Via r/LocalLlama
Tuesday, Jun 9, 2026 · 8:17AM
Summary

Hi all, I'm trying to wrap my head around hardware variables when it comes to LLM, and I have another question: what would be the cheapest way to run a 120B dense LLM at >10 tok/sec? I'm fine with Q5, ideally Q6 though. My goal would be advanced roleplay for RPG campaigns, and I need the answers to

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories