arxivcs.LGcs.AI2026-07-14
SteinGate: Tail-Sensitive Safe Reinforcement Learning via Stein Discrepancy
Yassine Chemingui, Chenhua Fan, Honghao Wei, Janardhan Rao Doppa
Safe reinforcement learning typically enforces safety by bounding expected cumulative costs, a criterion that often fails to detect rare but catastrophic tail events. To overcome these limitations, this paper introduces SteinGate, a boundary-aware distributional safety certificat…