Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Tried Qwen3.6-27B-UD-Q6_K_XL.gguf with CloudeCode, well I can't believe but it is usable

Via r/LocalLlama
Wednesday, Apr 22, 2026 · 9:44PM
Summary

So I tried to run Qwen3-27B-UD-Q6_K_XL.gguf with 200K context on my RTX 5090 using llama.cpp. I'm getting around 50 tok/s, which is fine I guess, I don't really know this stuff so it might be improvable. But what I want to say is, I haven't tried local models for coding for quite a long time, and he

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories