Human-assisted Robotic Policy Refinement via Action Preference Optimization
机构 * Gaoling School of Artificial Intelligence, Renmin University of China, Beijing(中国人民大学北京校区人工智能学院) ; Engineering Research Center of Next-Generation Intelligent Search(下一代智能搜索与推荐工程研究中心) ; Beijing Key Laboratory of Research on Large Models(北京大型模型研究重点实验室)
专题命中 后训练与偏好优化 :preference optimization(title,abstract);foundation model(abstract);分类 cs.AI
Comments Accepted By NeurIPS 2025