Muse Voice Transcribe Launched by MSL
First streaming audio model from Meta Superintelligence Lab does real-time ASR with diarization.
TLDR
Bowen Cheng announced the launch of Muse Voice Transcribe from Meta Superintelligence Lab. The model processes audio in 80ms chunks and decides at each step whether to keep listening or output text. It supports hour-long sessions, over 20 speakers, multilingual input with code-switching, and contextual biasing. Meta AI accounts and Mark Zuckerberg posted that it reaches the speed-accuracy pareto frontier on streaming transcription. The model is rolling out now to developers and appears in the Meta AI macOS app.
Combined views
323.8K
31 Sources, first seen 29d ago