One Video Teaches Robots 10-Minute Tasks: Skild AI Releases S1 Foundation Model, Breaking the Robot Training Paradigm
Skild AI has released the S1 robot foundation model, which learns multi-step physical tasks of up to 10 minutes from a single human demonstration video—without fine-tuning or post-training—achieving a 66% success rate on unseen tasks versus 9% for language-prompted VLA models. The approach treats video demonstrations as programs, leveraging trillion-scale simulation pretraining for in-the-wild generalization.