Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Agent recommendations

Via r/LocalLlama
Monday, Jun 22, 2026 · 2:23AM
Summary

Hi, I have a Strix Halo with 128GB setup that runs a couple of models (GPT-OSS 120b, Qwen3.5-122b, Gemma-4-31b) on llama-swap. GPT and Qwen run quite fast at 40-50T/s, while Gemma is a slow 4-5T/s but seems to have the best quality. I'd like to vibe code a personal Webproject in Python, using Pychar

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories