Parse-Augment-Distill: Learning Generalizable Bimanual Visuomotor Policies from Single Human Video

Open in new window