Theoretical Foundations and Effective Algorithms for Policy-Aware Simulator Learning
DGX agentarXiv:2605.29032v1 Announce Type: new Abstract: Model-based reinforcement learning (MBRL) agents typically learn world models by minimizing predictive loss. However, powerful RL optimizers inevitably