Alexandr Wang Rolls Out Muse Voice Transcribe
Meta executive releases first real-time audio perception model for streaming speech-to-text.
TLDR
Alexandr Wang, Chief AI Officer at Meta, posted that the company is rolling out Muse Voice Transcribe. He stated the model is the first real-time audio perception system from the team and reaches state-of-the-art results in streaming speech-to-text. Wang added that it performs speaker diarization and endpointing natively inside one model. Machine learning engineer Rohan Paul replied to the post, calling the release incredible and saying the world needs robust voice transcription models.
Combined views
156.1K
2 Sources, first seen 29d ago
