PEARL: Plan Exploration and Adaptive Reinforcement Learning for Multihop Tool Use
PEARL:多跳工具使用中的计划探索与自适应强化学习
Qihao Wang, Mingzhe Lu, Jiayue Wu, Yue Hu, Yanbing Liu
机构
*
Institute of Information Engineering, China(信息工程研究所,中国)
;
School of Cyber Security, University of Chinese Academy of Sciences, China(中国科学院大学网络安全学院,中国)
机构
*
Department of Materials Science and Engineering, University of California, Berkeley, California, United States(加州大学伯克利分校材料科学与工程系)
;
Materials Sciences Division, Lawrence Berkeley National Laboratory, California, United States(劳伦斯伯克利国家实验室材料科学部)
;
Laboratory of Artificial Chemical Intelligence (LIAC), Institute of Chemical Sciences and Engineering, École Polytechnique Fédérale de Lausanne (EPFL), Lausanne, Switzerland(日内瓦联邦理工学院化学科学与工程研究所人工化学智能实验室)
;
National Centre of Competence in Research (NCCR) Catalysis, École Polytechnique Fédérale de Lausanne (EPFL), Lausanne, Switzerland(日内瓦联邦理工学院催化研究国家中心)
Taxonomy of the Retrieval System Framework: Pitfalls and Paradigms
检索系统框架的分类:陷阱与范式
Deep Shah, Sanket Badhe, Nehal Kathrotia
机构
*
National Institute of Standards(国家标准技术研究所)
;
Department of Physics, Colorado State University, Fort Collins, CO 80523 USA(科罗拉多州立大学物理系)
;
Electrical Engineering Department, University of Colorado, Boulder, CO 80309 USA(科罗拉多大学电气工程系)
;
Google LLC, CA 94043 USA(谷歌公司)