Optimal Transport-Based Token Weighting scheme for Enhanced Preference Optimization
机构 * Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学人工智能学院) ; LLM Team, Shopee Pte. Ltd.(Shopee Pte. Ltd. 语言模型团队) ; Beijing Key Laboratory of Research on Large Models and Intelligent Governance and Engineering Research Center of Next-Generation Intelligent Search and Recommendation, MOE(北京大型模型与智能治理研究重点实验室)
Comments 24 pages, 11 figures. Accepted by ACL 2025 (main)