Skild AI’s S1 Robot Learns a New Task From One Video

Skild AI’s S1 robot model takes a video demonstration as its prompt instead of text and runs unseen tasks up to ten minutes long with no fine-tuning, averaging 66% success against 9% for a language-prompted policy.
artificial-intelligence
Author

Kabui, Charles

Published

2026-09-14

Keywords

robotics, in-context-learning, robot-foundation-model, video-demonstration, embodied-ai