Best AI News โ€” Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Why is opencode so slow in processing the prompt with llama server?

Via r/LocalLlama
Monday, May 11, 2026 ยท 11:40AM
Summary

I'm running opencode and llama-server locally. I have 32gb ram and 780m igpu. With Qwen3.6 I get around 21 t/s. Which should be decent but opencode just takes too long to process every input. What is it doing exactly? Tmux shows the available ram at the bottom (8+ GB available). Server startup comma

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories