Skip to content
VibekollenBETAVibekollen
BlogNVIDIA

Skild AI Taps NVIDIA Physical AI to Teach Robots New Tasks From a Single Video

Article image or reusable cover for NVIDIA

Skild AI has launched S1, a robot model that can learn new tasks from just a video without requiring retraining.

Instead of reprogramming robots for each new task, an operator can simply record a video of what needs to be done, and the model understands and performs it. S1 handles complex tasks lasting up to 10 minutes — such as plant potting, food preparation, and assembly work — and succeeds approximately 66 percent of the time on each step, compared to 9 percent for similar systems. Skild AI already has 60+ deployment partnerships and reached 100 million dollars in annual revenue 10 months after its first commercial deployment.

S1 takes a different approach: An operator records a video of the desired task and provides it to the model as a prompt. It interprets the demonstrated intent, objects and sequence, then maps them into actions for the robot in front of it — with no retraining — and often for a task not covered by its pretraining datase
Verbatim from the article at NVIDIA
Read the full story at NVIDIA →

Vibekollen prepared this summary with AI from the original publication. The content belongs to NVIDIA.

More to read