This feels like a genuine GPT-3 moment for robotics.
S1 can observe one video of a human completing a task it has never seen before, then carry out that long-horizon workflow without task-specific fine-tuning or post-training.
Moving from “collect data, train, and deploy” toward “show the robot once and let it act” would fundamentally change how quickly robots can adapt to new work.
Incredible progress from the
@SkildAI team.