Learning Human-Intention Priors from Large-Scale Human Demonstrations for Robotic Manipulation
DGX agentarXiv:2604.24681v1 Announce Type: new Abstract: Human videos contain rich manipulation priors, but using them for robot learning remains difficult because raw observations entangle scene understanding