A Unified Bellman Operator for Safety-Critical Reinforcement Learning
This study introduces a unified Bellman operator to maximize task performance in reinforcement learning while strictly adhering to safety constraints.
This study introduces a unified Bellman operator to maximize task performance in reinforcement learning while strictly adhering to safety constraints.