Exploiting Exogenous Structure for Sample-Efficient Reinforcement Learning
arXiv:2409.14557v4 Announce Type: replace-cross Abstract: We study a structured class of Markov Decision Processes, known as Exo-MDPs, in which the state space is partitioned into exogenous and endoge