parallelquant
September 6, 2026 · The Decoder

Meta releases always-on real-time transcription model

Meta's Superintelligence Labs released Muse Voice Transcribe, a real-time speech transcription model that processes audio in 80-millisecond chunks, distinguishes between speakers, and detects sentence boundaries. According to Artificial Analysis, it offers the most accurate streaming transcription at the lowest price currently on the market.

Why it matters: Meta explicitly frames this as infrastructure for AI agents that "listen in on real conversations" through devices like its camera glasses, signaling where its hardware-plus-AI strategy is headed. Cheap, low-latency, speaker-aware transcription is also a prerequisite for the "always-listening assistant" products several vendors have been previewing.

Related updates