• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Announcement

    Ditto-Bench tests GPT-6 Astra and MolmoAct2 on tougher robot-control tasks

    Its creators say the benchmark pairs simple goals with challenging physics to test both models beyond simple robot-control tasks.

    5 Sources, 1h ago, first seen 1h ago

    TLDR

    Ditto-Bench’s creators say GPT-6 Astra has shown impressive results on simple robot-control tasks. They built the benchmark to test Astra and MolmoAct2 on harder physics, including geometric constraints, precise contact and improvised tool use. They also examined design choices that may help the agents succeed or leave them stuck, and explored possibilities such as in-context learning for robotics.

    Combined views

    —

    5 Sources, first seen 1h ago

    Combined views

    —

    5 Sources, first seen 1h ago

    — likes
    — likes
    — comments
    — saves
    — reposts
    — comments
    — saves
    — reposts
    Featured Source

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    5 Sources

    Lin Long@Kylin_0LongGPT-6 Astra has shown impressive results in simple robot control tasks. But what if the tasks become more challenging and involve complex geometric constraints, precise contact, or improvised tool use? To find out, we: - Built Ditto-Bench, where simple goals meet challenging physics, and tested both Astra and MolmoAct2. - Systematically ablated key design choices to understand what makes these robot-control agents succeed, or get stuck. - Explored what new possibilities these models could open up for robotics, such as in-context learning. More details 🧵👇1h
    Jaemin Cho@jmin__choRT @Kylin_0Long: GPT-6 Astra has shown impressive results in simple robot control tasks. But what if the tasks become more challenging and…1h
    Han Lin@hanlin_hlGreat and timely analysis for GPT-6 Astra as robotics control policy!29m

    5 Sources

    Lin Long@Kylin_0LongGPT-6 Astra has shown impressive results in simple robot control tasks. But what if the tasks become more challenging and involve complex geometric constraints, precise contact, or improvised tool use? To find out, we: - Built Ditto-Bench, where simple goals meet challenging physics, and tested both Astra and MolmoAct2. - Systematically ablated key design choices to understand what makes these robot-control agents succeed, or get stuck. - Explored what new possibilities these models could open up for robotics, such as in-context learning. More details 🧵👇1h
    Jaemin Cho@jmin__choRT @Kylin_0Long: GPT-6 Astra has shown impressive results in simple robot control tasks. But what if the tasks become more challenging and…1h
    Han Lin@hanlin_hlGreat and timely analysis for GPT-6 Astra as robotics control policy!29m