S1 is Skild AI’s in-context robotic foundation model. It learns a new manipulation task from one visual demonstration and executes it without task-specific fine-tuning or post-training.
About
S1 is the named flagship in-context-learning model within Skild AI’s broader Skild Brain platform. A user provides one video demonstration, and S1 conditions its robot policy on that example to perform the task without updating model weights. Skild demonstrates the approach on seen and unseen manipulation tasks including plant potting, pancake preparation, pour-over coffee and kit assembly.
The company reports unseen tasks lasting up to 10 minutes and a sevenfold improvement over language-only prompting on its evaluation set. Its plant-potting example spans an 11-minute demonstration-to-execution workflow. Skild says S1 is already operating with commercial partners, while public model weights, code and training data have not been released.