Alexandr Wang Rolls Out Muse Voice Transcribe
Meta executive releases first real-time audio perception model for streaming speech-to-text.
Alexandr Wang, Chief AI Officer at Meta, posted that the company is rolling out Muse Voice Transcribe. He stated the model is the first real-time audio perception system from the team and reaches state-of-the-art results in streaming speech-to-text. Wang added that it performs speaker diarization and endpointing natively inside one model. Machine learning engineer Rohan Paul replied to the post, calling the release incredible and saying the world needs robust voice transcription models.
1/ today we're rolling out muse voice transcribe, our first real-time audio perception model - SOTA in streaming speech-to-text. also handles speaker diarization and endpointing natively in a single model.
Alexandr Wang Rolls Out Muse Voice Transcribe
Meta executive releases first real-time audio perception model for streaming speech-to-text.
Alexandr Wang, Chief AI Officer at Meta, posted that the company is rolling out Muse Voice Transcribe. He stated the model is the first real-time audio perception system from the team and reaches state-of-the-art results in streaming speech-to-text. Wang added that it performs speaker diarization and endpointing natively inside one model. Machine learning engineer Rohan Paul replied to the post, calling the release incredible and saying the world needs robust voice transcription models.
1/ today we're rolling out muse voice transcribe, our first real-time audio perception model - SOTA in streaming speech-to-text. also handles speaker diarization and endpointing natively in a single model.
