By combining visual demonstrations with agentic code refinement and object-centric representations, RAPID enables robots to learn generalizable manipulation skills that transfer across object variations and scene configurations.
RAPID automatically generates reusable robot programs from a single human video demonstration. It uses AI coding agents to create, test, and refine programs that work across different objects and environments by learning the underlying strategy rather than memorizing specific motions.