LLMs handle speech well once you run speech-to-text. They don't hear the rest: a bird outside, a glass breaking two rooms away, a smoke alarm two floors down. I've been working on an experimental open-source framework that aims to close that gap: continuous audio in, event-gated recognition out, a c