Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Struggling with Qwen3.6 27B / 35B locally (3090) slow responses, breaking code looking for better setup + auto model switching

Via r/LocalLlama
Tuesday, May 5, 2026 · 8:51AM
Summary

Hey everyone, I’ve been experimenting with running Qwen models locally on my setup: GPU: RTX 3090 (24GB VRAM) RAM: 64GB CPU: Ryzen 5700X OS: Windows 11 What I’m currently running Qwen 3.6 35B (UD Q4_K_M) llama-server.exe -m "C:\Users\Dino\.lmstudio\models\unsloth\Qwen3.6-35B-A3B-GGUF\Qwen3.6-35B-A3B

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories