The Irrational Machine: Neurosis and the Limits of Algorithmic Safety
Daniel Howard
机构
*
Howard Science Limited, Malvern, UK(霍华德科学有限公司,英国马尔文)
;
QinetiQ Fellow, UK(QinetiQ Fellow,英国)
;
Member of Senior Common Room, Pembroke College, University of Oxford(奥克斯福德大学彭伯里学院高级共同房间成员)
机构
*
University of Tübingen, Zuse School ELIZA(图宾根大学Zuse学校ELIZA)
;
University of Tübingen, Tübingen AI Center(图宾根大学图宾根人工智能中心)
;
University of Tübingen, Tübingen AI Center, MPI for Informatics, SIC(图宾根大学图宾根人工智能中心、马克斯·普朗克信息研究所、SIC)
Trust Region Reward Optimization and Proximal Inverse Reward Optimization Algorithm
Yang Chen, Menglin Zou, Jiaqi Zhang, Yitan Zhang, Junyi Yang, Gael Gendron, Libo Zhang, Jiamou Liu, Michael J. Witbrock
机构
*
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
University of Auckland(奥克兰大学)
;
Chongqing University(重庆大学)
专题命中
模仿学习与强化学习
:robotics(abstract);分类 cs.AI、cs.LG
CommentsAccepted to NeurIPS 2025. Title used at submission and review: PIRO: Toward Stable Reward Learning for Inverse RL via Monotonic Policy Divergence Reduction
机构
*
Graduate School of Informatics, Nagoya University, Japan(名古屋大学信息学研究科)
;
RIKEN Center for Advanced Intelligence Project, Japan(RIKEN高级智能项目研究中心)
;
Graduate School of Arts and Sciences, The University of Tokyo, Japan(东京大学文学系研究科)
;
Project team for SIP, Japan(SIP项目团队)
;
Japan Agency for Marine-Earth Science and Technology, Japan(日本海洋地球科学技术机构)
;
Faculty of Education, Shitennoji University, Japan(世田谷大学教育学部)
;
Graduate School of Engineering, The University of Tokyo, Japan(东京大学工学研究科)
;
Graduate School of Science and Technology, Niigata University, Japan(新潟大学科学技术研究科)
;
Graduate School of Science, Nagoya University, Japan(名古屋大学理学研究科)
;
Principles of Informatics Research Division, National Institute of Informatics, Japan(信息学原理研究部门,日本信息处理技术研究所)
;
Graduate School of Information Science, The University of Osaka, Japan(大阪大学信息科学研究科)
机构
*
Wuhan University(武汉大学)
;
DAMO Academy, Alibaba Group(达摩院,阿里巴巴集团)
;
Hupan Lab(虎扑实验室)
;
The Chinese University of Hong Kong(香港中文大学)
;
Tsinghua University(清华大学)
;
Huazhong University of Science and Technology(华中科技大学)
;
Zhejiang University(浙江大学)
MLLM-Fabric: Multimodal Large Language Model-Driven Robotic Framework for Fabric Sorting and Selection
Liman Wang, Hanyang Zhong, Tianyuan Wang, Shan Luo, Jihong Zhu
机构
*
School of Physics, Engineering and Technology, University of York(物理、工程与技术学院,约克大学)
;
Department of Engineering, King’s College London(工程学院,伦敦国王学院)
CommentsThe work consists of three chapters, includes 12 figures, 4 tables, 31 references, and 1 appendix. A version of this work has been accepted for presentation at the 2025 IEEE 8th International Conference on Methods and Systems of Navigation and Motion Control