GPT-6 Astra reportedly learns robot tasks from video without a text prompt
The user describes a robotics demo with a different environment, camera angle and layout, and says Astra chooses between end-effector and joint-space control.
TLDR
A user describes GPT-6 Astra performing in-context learning for mobile manipulation: inferring what to do from video without a text prompt. They say it handles a different environment, camera angle and layout, and chooses between directing the robot’s working end and controlling its joints. Another user sharing the demo calls the development “amazing for deployment speed” and “frustrating for foundational models.”
Combined views
66K
5 Sources, first seen 20d ago