World Wires · story 31284 · corroborated · 1 source(s)

A Unified Bellman Operator for Safety-Critical Reinforcement Learning

This study introduces a unified Bellman operator to maximize task performance in reinforcement learning while strictly adhering to safety constraints.

Open in the desk

Coverage

What this site indexes