Skild AI Shows S1 Model Learning 10-Minute Tasks From One Video
Skild AI's S1 model performs long-horizon tasks never seen in pre-training from single video prompts.
Skild AI founder Deepak Pathak posted that in-context learning for robotics has arrived with the S1 model. It handles tasks over 10 minutes long that were never seen during pre-training, prompted by one video example and without fine-tuning. Investor Alfred Lin at Sequoia called the single-prompt execution of long-horizon tasks a game changer. Skild researcher Roozbeh Mottaghi noted the absence of test-task leakage and the model's ability to manage extended independent execution. Other observers including Chris Paxton and Sequoia Capital highlighted the demonstration as progress toward general-purpose robot foundation models.
Introducing S1, our new foundation model that learns from one example. It can be taught 10-minute long tasks that it has never seen before, from one video prompt without any fine-tuning. Watch S1 operate in real-time via in-context learning:


