A KL-regularization Framework for Learning to Plan with Adaptive Priors
DGX agentarXiv:2510.04280v2 Announce Type: replace-cross Abstract: Effective exploration remains a central challenge in model-based reinforcement learning (MBRL), particularly in high-dimensional continuous co