Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Another shout out to llama.cpp build b9455 2x3090

Via r/LocalLlama
Wednesday, Jun 3, 2026 · 5:05AM
Summary

https://preview.redd.it/xyvtkzwr005h1.png?width=645&format=png&auto=webp&s=aebd5b5ef79255247c9bc91fb69d8423a0c61f86 As you guys know, the next highest quant is Unsloth's /Qwen3.6-27B-UD-Q8_K_XL.gguf. With llama.cpp before, i was getting 30-50 tk/s. vllm was kicking llama's ass with its tensor splits

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories