Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Everyone here self-hosts inference. Almost nobody self-hosts the tooling around it. That feels backwards to me.

Via r/LocalLlama
Saturday, May 30, 2026 · 9:48PM
Summary

Been running local models for a while, currently a 3090 box for the daily driver stuff and a second machine with a pair of older cards I use for batch jobs. Nothing exotic by this sub's standards. Standard reasons: cost at volume, not sending data to an API, and honestly just liking that it's mine.

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories