Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Qwen 35b a3b surprises me

Via r/LocalLlama
Monday, May 18, 2026 · 3:50PM
Summary

Just wanted to share that I'm pretty happy about Qwen 35b a3b agentic coding performance. I'm running the model in q80 quant, kv cache both q8_0 as well, with 262144 in 4090 + 5060 ti, via llama.cpp backend with claude code pointing to localhost. For demo/data analytics purposes, it works pretty wel

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories