ConTraIRL: Factorized Contrastive Abstractions for Transferable IRL
DGX agentarXiv:2606.03017v1 Announce Type: cross Abstract: Reward transfer in Inverse Reinforcement Learning (IRL) is unreliable when policies must generalize to unseen combinations of environment dynamics and