• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Report

    Utopai X debuts at No. 2 on Artificial Analysis’s text-to-video-with-audio leaderboard

    Artificial Analysis ranks the new Utopai Studios model behind Wan 3.0 overall, but first for audio synchronization and physics in its capability rankings.

    AA
    SA
    4 Sources, ,

    TLDR

    Artificial Analysis says Utopai Studios launched Utopai X, a video model post-trained on MiniMax H3, in its PAI platform on September 30. It ranks the model second behind Wan 3.0 on its refreshed Text to Video with Audio leaderboard and first for audio synchronization and physics. Utopai X costs 93 PAI credits per second of video and has no public API.

    Combined views

    52.8K

    4 Sources, first seen 11h ago

    Combined views

    52.8K

    4 Sources, first seen 11h ago

    291 likes
    11h ago
    first seen 11h ago
    291 likes
    24 comments
    110 saves
    24 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Featured Source
    24 comments
    110 saves
    24 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    4 Sources

    @ArtificialAnlysUtopai X (based on MiniMax H3) debuts at #2 on the Artificial Analysis Text to Video Leaderboard, behind only Wan 3.0 Utopai X is the new video model from Utopai Studios, post-trained on MiniMax H3 and launching today inside PAI, Utopai’s cinematic storytelling web platform. Utopai Studios is an AI-native film and TV studio, with three wide theatrical release films and two series scheduled for release in 2027. In AA-Video-T2V v2.0, our refreshed Text to Video with Audio leaderboard, Utopai X ranks #2, behind Wan 3.0 and ahead of Dreamina Seedance 2.5. As of today, it is the highest ranked of the three models on the leaderboard built on MiniMax H3, ahead of MiniMax H3 itself and fal's MiniMax H3 Max. Utopai X is available in PAI at 93 PAI credits per second of video, with no public API. PAI plans start at $15 per month (billed monthly) for 1,000 credits. Congratulations to @UtopaiStudios on the release! See below for our analysis and example outputs of Utopai X in the Artificial Analysis Video Arena 🧵
    @svpinoThis is the #2 model on the Artificial Analysis Text-to-Video Leaderboard with Audio worldwide. No other US-based company has a model higher than this. I checked their website, and their model's production quality is really high. They use it for sports, film, and all sorts of TV projects.

    4 Sources

    @ArtificialAnlysUtopai X (based on MiniMax H3) debuts at #2 on the Artificial Analysis Text to Video Leaderboard, behind only Wan 3.0 Utopai X is the new video model from Utopai Studios, post-trained on MiniMax H3 and launching today inside PAI, Utopai’s cinematic storytelling web platform. Utopai Studios is an AI-native film and TV studio, with three wide theatrical release films and two series scheduled for release in 2027. In AA-Video-T2V v2.0, our refreshed Text to Video with Audio leaderboard, Utopai X ranks #2, behind Wan 3.0 and ahead of Dreamina Seedance 2.5. As of today, it is the highest ranked of the three models on the leaderboard built on MiniMax H3, ahead of MiniMax H3 itself and fal's MiniMax H3 Max. Utopai X is available in PAI at 93 PAI credits per second of video, with no public API. PAI plans start at $15 per month (billed monthly) for 1,000 credits. Congratulations to @UtopaiStudios on the release! See below for our analysis and example outputs of Utopai X in the Artificial Analysis Video Arena 🧵
    @svpinoThis is the #2 model on the Artificial Analysis Text-to-Video Leaderboard with Audio worldwide. No other US-based company has a model higher than this. I checked their website, and their model's production quality is really high. They use it for sports, film, and all sorts of TV projects.