Embedding Safety into RL: A New Take on Trust Region Methods
机构 * Max Planck Institute for Human Cognitive and Brain Sciences(马克斯·普朗克人类认知与脑科学研究所) ; Center for Scalable Data Analytics and Artificial Intelligence(可扩展数据分析与人工智能中心)
专题命中 安全训练 :safety(title,abstract);分类 cs.LG
Comments Accepted at ICML 2025
Journal ref Proceedings of the 42nd International Conference on Machine Learning, Vancouver, Canada. PMLR 267, 2025