Controlling Output Rankings in Generative Engines for LLM-based Search
在基于大语言模型的搜索生成引擎中控制输出排名
Haibo Jin, Ruoxi Chen, Peiyan Zhang, Yifeng Luo, Huimin Zeng, Man Luo, Haohan Wang
机构
*
School of Information Sciences, University of Illinois at Urbana-Champaign, IL, USA(伊利诺伊大学厄巴纳-香槟分校信息科学学院)
;
Independent Researcher, Starc Institute(Starc研究所独立研究者)
;
Research Scientist, Intel Labs, Santa Clara, CA, USA(英特尔实验室研究科学家)
机构
*
Department of Statistics and Data Science, Southern University of Science and Technology, China(统计与数据科学系,南方科技大学)
;
Department of Statistics, University of California at Riverside, USA(统计系,加州大学河滨分校)
;
College of Computing and Data Science, Nanyang Technological University, Singapore(计算与数据科学学院,南洋理工大学)
;
School of Artificial Intelligence, The Chinese University of Hong Kong, Shenzhen, China(人工智能学院,香港中文大学(深圳))
Scaling Offline Model-Based RL via Jointly-Optimized World-Action Model Pretraining
通过联合优化的世界-动作模型扩展离线模型基于的强化学习
Jie Cheng, Ruixi Qiao, Yingwei Ma, Binhua Li, Gang Xiong, Qinghai Miao, Yongbin Li, Yisheng Lv
机构
*
State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Alibaba Group(阿里巴巴集团)
机构
*
Department of Computer Science & Information Systems, Birla Institute of Technology and Science, Pilani, India(印度比拉理工学院计算机科学与信息系统系)
;
Birdeye Inc.(Birdeye公司)
Paying Less Generalization Tax: A Cross-Domain Generalization Study of RL Training for LLM Agents
支付更少的泛化税:RL训练对LLM代理跨域泛化能力的研究
Zhihan Liu, Lin Guan, Yixin Nie, Kai Zhang, Zhuoqun Hao, Lin Chen, Asli Celikyilmaz, Zhaoran Wang, Na Zhang
机构
*
Meta Superintelligence Labs(Meta超智能实验室)
;
FAIR at Meta(Meta的FAIR)
;
Northwestern University(西北大学)
;
The Ohio State University(俄亥俄州立大学)
;
University of Pennsylvania(宾夕法尼亚大学)
Empowering LLMs for Structure-Based Drug Design via Exploration-Augmented Latent Inference
通过探索增强的潜在推理增强LLM用于基于结构的药物设计
Xuanning Hu, Anchen Li, Qianli Xing, Jinglong Ji, Hao Tuo, Bo Yang
机构
*
College of Computer Science and Technology, Jilin University(吉林大学计算机科学与技术学院)
;
Key Laboratory of Symbolic Computation and Knowledge Engineering of Ministry of Education, Jilin University(教育部符号计算与知识工程重点实验室)
;
Department of Computer Science, Aalto University(艾尔沃斯大学计算机科学系)
;
College of Artificial Intelligence, Jilin University(吉林大学人工智能学院)
Comments14 pages, 4 figures, proceedings of the International Workshop on Active Inference 2025. Erratum v1: in Eq. (50), $p(y_t, Θ, u_t \mid y_{*}, \mathcal{D}_k)$ should have been $p(y_t, Θ\mid u_t, y_{*}, \mathcal{D}_k)$