Jul 2026
Directional Constraints for Efficient Exploration in Safe Reinforcement Learning
This work proposes an extension of the ATACOM framework, a state-of-the-art reliable safety layer that can be integrated with existing Reinforcement Learning algorithms to enforce constraints derived from prior knowledge of the system or learned directly from data.
Paolo Magliano, Puze Liu, Jan Peters et al.
· arXiv.org · 0 citations