Best AI News โ€” Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

[audio.cpp] What Does the Fox Say: 4 ASR models (Nemotron 3.5 ASR, Higgs Audio STT, VibeVoice ASR, and Hviske ASR) in native C++/GGML, init streaming support, and 327s of audio transcribed in 2.17s.

Via r/LocalLlama
Thursday, Jul 9, 2026 ยท 2:17AM
Summary

I just pushed a new audio.cpp update with streaming support and 4 ASR/STT models: Nemotron 3.5 ASR, Higgs Audio STT, VibeVoice ASR, and Hviske ASR (da only). Overall 1.07x to 2.41x faster than Python. I decided to drop Parakeet-TDT since good implementations already exist, and I find the Nemotron mo

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories