机构
*
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
Southwestern University of Finance and Economics(西南财经大学)
Layer-Order Inversion: Rethinking Latent Multi-Hop Reasoning in Large Language Models
层序倒置:重新思考大语言模型中的潜在多跳推理
Xukai Liu, Ye Liu, Jipeng Zhang, Yanghai Zhang, Kai Zhang, Qi Liu
机构
*
State Key Laboratory of Cognitive Intelligence(认知智能国家重点实验室)
;
University of Science and Technology of China(中国科学技术大学)
;
The Hong Kong University of Science and Technology(香港科技大学)
Guangyu Wang, Jingkun Yue, Siqi Zhang, Yu Liu, Xiaoyu Wang, Mingyuan Meng, Changwei Ji, Zongbo Han, Yulin Wang, Yang Yue, Frank Fu, Ting Chen, Song Wu, Ziwei Liu, Jiangning Song, Ming Li, Gao Huang, Xiaohong Liu, Athanasios Vasilakos, Xingcai Zhang, Ping Zhang, Yong Li
机构
*
State Key Laboratory of Networking and Switching Technology, Beijing University of Posts and Telecommunications, Beijing, China(网络与交换技术国家重点实验室,北京邮电大学,北京,中国)
;
Department of Engineering Science, University of Oxford, Oxford, United Kingdom(英国牛津大学工程科学系,牛津,英国)
;
Institute of Medical Artificial Intelligence, South China Hospital, Medical School, Shenzhen University, Shenzhen, Guangdong, China(医学人工智能研究所,南方医院,医学学院,深圳大学,深圳,广东,中国)
;
Zhongguancun Academy & Zhongguancun Institute of Artificial Intelligence, Beijing, China(中关村学院及中关村人工智能研究院,北京,中国)
;
Beijing National Research Center for Information Science and Technology (BNRist), Tsinghua University, 100084, Beijing, China(北京信息科学与技术国家研究中心(BNRist),清华大学,100084,北京,中国)
;
Department of Chemical and Nano Engineering, University of California, San Diego, La Jolla, CA, USA(美国加州大学圣地亚哥分校化学与纳米工程系,La Jolla,CA,美国)
;
Nanyang Technological University, Singapore(新加坡南洋理工大学)
;
Monash Biomedicine Discovery Institute and Department of Biochemistry and Molecular Biology, Monash University, Melbourne, Victoria, Australia(莫纳什大学生物医学发现研究所和生物化学与分子生物学系,墨尔本,维多利亚,澳大利亚)
;
David R. Cheriton School of Computer Science, University of Waterloo, Waterloo, Ontario, Canada(加拿大滑铁卢大学戴维·R·切里顿计算机科学学校,滑铁卢,安大略,加拿大)
;
Department of ICT and Center for AI Research, University of Agder (UiA), Jon Lilletuns vei 9, Grimstad, Norway(挪威阿格德大学(UiA)信息与通信技术系及人工智能研究中心,Jon Lilletuns vei 9,Grimstad,挪威)
;
Department of Electronic Engineering, Tsinghua University, Beijing, China(清华大学电子工程系,北京,中国)
Comments30 pages, 3 figures, 2 tables. This paper includes a large amount of work that was done subsequently after comments and supersedes a previous paper submitted to arXiv (Reference number: 2403.16906 (https://arxiv.org/abs/2403.16906). I have not deleted or replaced the latter in case the moderators prefer both papers to be readable side-by-side
Interpreting Protein Language Model Embeddings via Orthogonal Projection for Protein Fitness Prediction
通过正交投影解释蛋白质语言模型嵌入以用于蛋白质适应性预测
Paulo Yanez Sarmiento, Pia Francesca Rissom, Manuel Pfeuffer, Marco Simnacher, Jordan F. Safer, Sumaiya Iqbal, Henrike O. Heyne, Nadja Klein, Bernhard Y. Renard
机构
*
Hasso Plattner Institute, University of Potsdam(波茨坦大学哈索·普拉特纳研究所)
;
Broad Institute of MIT and Harvard(麻省理工学院及哈佛大学布罗德研究所)
;
Humboldt University of Berlin(柏林洪堡大学)
;
Karlsruhe Institute of Technology(卡尔斯鲁厄理工学院)