LIMSSR: LLM-Driven Sequence-to-Score Reasoning under Training-Time Incomplete Multimodal Observations
LIMSSR:基于大语言模型的训练时不完整多模态观察下的序列到评分推理
Huangbiao Xu, Huanqi Wu, Xiao Ke, Yuxin Peng
机构
*
Fujian Provincial Key Laboratory of Networking Computing(网络计算与智能信息处理福建省重点实验室)
;
Intelligent Information Processing, College of Computer(智能信息处理学院)
;
Data Science, Fuzhou University(数据科学,福州大学)
;
Engineering Research Center of Big Data Intelligence, Ministry of Education(大数据智能工程研究中心,教育部)
;
Wangxuan Institute of Computer Technology, Peking University(王宣计算机技术研究所,北京大学)
Outbidding and Outbluffing Elite Humans: Mastering Liar's Poker via Self-Play and Reinforcement Learning
超越精英人类:通过自我对战与强化学习掌握骗子扑克
Richard Dewey, Janos Botyanszki, Ciamac C. Moallemi, Andrew T. Zheng
机构
*
Allometry Labs(Allometry实验室)
;
Graduate School of Business, Columbia University(哥伦比亚大学商学院)
;
Sauder School of Business, University of British Columbia(不列颠哥伦比亚大学萨德尔商学院)