Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Gemma4 26b & E4B are crazy good, and replaced Qwen for me!

Via r/LocalLlama
Wednesday, Apr 15, 2026 · 7:56PM
Summary

My pre-gemma 4 setup was as follows: Llama-swap, open-webui, and Claude code router on 2 RTX 3090s + 1 P40 (My third 3090 died, RIP) and 128gb of system memory Qwen 3.5 4B for semantic routing to the following models, with n_cpu_moe where needed: Qwen 3.5 30b A3B Q8XL - For general chat, basic docum

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories