Best AI News โ€” Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

llama.cpp benchmark native vs. non native NVFP4 on Blackwell - summary

Via r/LocalLlama
Wednesday, Apr 29, 2026 ยท 12:27PM
Summary

I tested two llama.cpp builds on the same Qwen3.6-27B-NVFP4 model. llama-bench reports the model label as qwen35 27B NVFP4, but the actual tested model is Qwen3.6-27B-NVFP4. Test platform GPU: NVIDIA GeForce RTX 5090 CPU: AMD Ryzen 9 9950X3D RAM: 128 GB DDR5 5600 CL36 Backend: CUDA Tested builds b89

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories