Best AI News โ€” Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Acceptable prompt processing speed for you?

Via r/LocalLlama
Sunday, Apr 19, 2026 ยท 8:00AM
Summary

I am currently optimising some ancient hardware to run qwen3 (4xV100s) but the lack of flash attention means that at longer contexts the processing starts to really slow down. For agentic coding work what processing speeds and contexts lengths do you consider as acceptable or good?

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories