Reinforcement Learning with Discrete Diffusion Policies for Combinatorial Action Spaces
DGX agentarXiv:2509.22963v3 Announce Type: replace Abstract: Reinforcement learning (RL) struggles to scale to large, combinatorial action spaces common in many real-world problems. This paper introduces a nov