• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Atlas user describes generating camera-controlled video from a few scene views

    The user says Atlas combines input images, previously generated views and camera positions and orientations into a “spatial context” that guides future frames.

    WD
    ML
    BM
    4 Sources, ,

    TLDR

    A user describes feeding Atlas a few views of the “gardenvase” scene, turning the camera to face the other side and using text to prompt a fourth view containing an Atlas ad. They say they can then generate a smooth, camera-controlled video sequence through the scene. According to the same user, Atlas grounds future frames in a “spatial context” combining input images, its own generated views and their camera positions and orientations.

    Combined views

    14.5K

    4 Sources, first seen 29d ago

    Combined views

    14.5K

    4 Sources, first seen 29d ago

    225 likes
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    29d ago
    first seen 29d ago
    225 likes
    17 comments
    37 saves
    3.5K reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    17 comments
    37 saves
    3.5K reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    4 Sources

    @BenMildenhallAtlas aggregates your input images, its own previously generated views, and their respective camera poses into a “spatial context” that grounds all future frames and imbues the model with a strong understanding of the 3D world and multiview geometry. https://worldlabs.ai/blog/atlas
    @willdepuethere is no such thing as a world model, there are only autoregressive video models and mistakes
    @ManlingLi_RT @theworldlabs: Introducing Atlas: The world's first multimodal world model that generates image and video frames with pixel-perfect cam…

    4 Sources

    @BenMildenhallAtlas aggregates your input images, its own previously generated views, and their respective camera poses into a “spatial context” that grounds all future frames and imbues the model with a strong understanding of the 3D world and multiview geometry. https://worldlabs.ai/blog/atlas
    @willdepuethere is no such thing as a world model, there are only autoregressive video models and mistakes
    @ManlingLi_RT @theworldlabs: Introducing Atlas: The world's first multimodal world model that generates image and video frames with pixel-perfect cam…