GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation
arXiv:2608.03753v1 Announce Type: new Abstract: Learning long-horizon manipulation skills with reinforcement learning remains challenging due to the complexity of reward design, the limited guidance o