2025
ProSh: Probabilistic Shielding for Model-free Reinforcement Learning
not assume access to a simulator or environment abstraction, and Safety is a major concern in reinforcement learning (RL): we aim at is compatible with continuous environments. developing RL systems that not only perform optimally, but are also We also show that optimizing over s...