Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Hy3 (295B MoE) and NVIDIA Nemotron-Labs-Audex-30B-A3B (audio-capable 30B MoE) GGUF quants

Via r/LocalLlama
Saturday, Jul 11, 2026 · 1:15AM
Summary

Sharing two GGUF quant sets, both with the same treatment: imatrix quantization, KLD/PPL measured against BF16 reference logits, llama-bench throughput numbers, and all raw benchmark data included in the repos. No vibes-based "quality tested" claims — every number is reproducible from the files in t

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories