A Reward-Free Viewpoint on Multi-Objective Reinforcement Learning
arXiv:2604.24532v1 Announce Type: new Abstract: Many sequential decision-making tasks involve optimizing multiple conflicting objectives, requiring policies that adapt to different user preferences. I