1
SteinGate: Tail-Sensitive Safe Reinforcement Learning via Stein Discrepancy
利用Stein散度提升安全强化学习对尾部风险的敏感度,UAI 2026收录新方法
arXiv:2607.13175v1 Announce Type: cross Abstract: Safe reinforcement learning typically enforces safety by bounding expected cumulative costs, a crite…