• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Alexandr Wang Rolls Out Muse Voice Transcribe

    Meta executive releases first real-time audio perception model for streaming speech-to-text.

    AW
    RP
    2 Sources, 29d ago, first seen 29d ago

    TLDR

    Alexandr Wang, Chief AI Officer at Meta, posted that the company is rolling out Muse Voice Transcribe. He stated the model is the first real-time audio perception system from the team and reaches state-of-the-art results in streaming speech-to-text. Wang added that it performs speaker diarization and endpointing natively inside one model. Machine learning engineer Rohan Paul replied to the post, calling the release incredible and saying the world needs robust voice transcription models.

    Combined views

    156.1K

    2 Sources, first seen 29d ago

    Combined views

    156.1K

    2 Sources, first seen 29d ago

    2K likes
    2K likes
    85 comments
    389 saves
    149 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Featured Source
    85 comments
    389 saves
    149 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    2 Sources

    @alexandr_wang1/ today we're rolling out muse voice transcribe, our first real-time audio perception model - SOTA in streaming speech-to-text. also handles speaker diarization and endpointing natively in a single model.
    @rohanpaul_ai@alexandr_wang Incredible release. the world needs robust voice transcription models.

    2 Sources

    @alexandr_wang1/ today we're rolling out muse voice transcribe, our first real-time audio perception model - SOTA in streaming speech-to-text. also handles speaker diarization and endpointing natively in a single model.
    @rohanpaul_ai@alexandr_wang Incredible release. the world needs robust voice transcription models.