HyPER: Bridging Exploration and Exploitation for Scalable LLM Reasoning with Hypothesis Path Expansion and Reduction
HyPER:通过假设路径扩展与缩减实现可扩展的大语言模型推理中的探索与利用平衡
Shengxuan Qiu, Haochen Huang, Shuzhang Zhong, Pengfei Zuo, Meng Li
机构
*
Institute for Artificial Intelligence, Peking University, Beijing(北京大学人工智能研究院)
;
Huawei(华为)
;
School of Integrated Circuits, Peking University, Beijing(北京大学集成电路学院)
Learning Graph Foundation Models on Riemannian Graph-of-Graphs
在黎曼图-图上学习图基础模型
Haokun Liu, Zezhong Ding, Xike Xie
机构
*
School of Biomedical Engineering, University of Science and Technology of China (USTC), Suzhou, Jiangsu, China(生物医学工程学院,中国科学技术大学(USTC),苏州,江苏,中国)
;
Data Darkness Lab, Suzhou Institute for Advanced Research, USTC, Suzhou, Jiangsu, China(Data Darkness实验室,苏州市先进研究院,USTC,苏州,江苏,中国)
;
School of Artificial Intelligence and Data Science, USTC, Hefei, Anhui, China(人工智能与数据科学学院,USTC,合肥,安徽,中国)
Turning Drift into Constraint: Robust Reasoning Alignment in Non-Stationary Multi-Stream Environments
将漂移转化为约束:非稳态多流环境中的鲁棒推理对齐
Xiaoyu Yang, En Yu, Wei Duan, Jie Lu
机构
*
Australian Artificial Intelligence Institute (AAII)(澳大利亚人工智能研究所)
;
Faulty of Engineering and Information Technology(工程与信息技术学院)
;
University of Technology Sydney(悉尼技术大学)
;
Australia(澳大利亚)
专题命中
推理与问题求解
:large language model(abstract);language model(abstract);preference optimization(abstract);分类 cs.AI、cs.LG
C-CoT: Counterfactual Chain-of-Thought with Vision-Language Models for Safe Autonomous Driving
C-CoT:基于视觉-语言模型的反事实链式推理用于安全自动驾驶
Kefei Tian, Yuansheng Lian, Kai Yang, Xiangdong Chen, Shen Li
机构
*
College of Transportation, Tongji University(同济大学交通运输学院)
;
Department of Civil Engineering, Tsinghua University(清华大学土木工程系)
;
School of Vehicle and Mobility, Tsinghua University(清华大学车辆与移动系统学院)
;
Department of Civil and Environmental Engineering, National University of Singapore(新加坡国立大学土木与环境工程系)
Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation
沙盒计划,开放世界导航:学习物理基础的抽象经验以实现具身导航
Zhixuan Shen, Jiawei Du, Ziyu Guo, Han Luo, Lilan Peng, Joey Tianyi Zhou, Haonan Luo, Tianrui Li
机构
*
School of Computing and Artificial Intelligence, Southwest Jiaotong University, China(计算机与人工智能学院,西南交通大学,中国)
;
Centre for Frontier AI Research A*STAR, Singapore(前沿人工智能研究A*STAR中心,新加坡)
;
School of Computer Science, University of Leeds, UK(计算机科学学院,利兹大学,英国)
HTPO: Towards Exploration-Exploitation Balanced Policy Optimization via Hierarchical Token-level Objective Control
HTPO: 通过分层令牌级目标控制实现探索-利用平衡的策略优化
Xincheng Yao, Ruoqi Li, Cheng Chen, Daoxin Zhang, Yi Wu, Yao Hu, Chongyang Zhang
机构
*
School of Information Science and Electronic Engineering, Shanghai Jiao Tong University(上海交通大学信息科学与电子工程学院)
;
Xiaohongshu Inc.(小红书公司)
;
MoE Key Lab of Artificial Intelligence, AI Institute, Shanghai Jiao Tong University(上海交通大学人工智能研究院)
专题命中
推理与问题求解
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG
机构
*
Department of XXX, University of YYY, Location, Country(XXX系,YYY大学,地点,国家)
;
School of ZZZ, Institute of WWW, Location, Country(ZZZ学院,WWW研究所,地点,国家)
;
JP Morgan AI Research, London, UK(摩根大通AI研究,伦敦,英国)
;
JP Morgan AI Research, New York, USA(摩根大通AI研究,纽约,美国)
专题命中
推理与问题求解
:large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
Effective Explanations Support Planning Under Uncertainty
有效解释支持在不确定性下的规划
Hanqi Zhou, Britt Besch, Charley M. Wu, Tobias Gerstenberg
机构
*
University of Tübingen(图宾根大学)
;
Technical University Darmstadt(达姆施塔特技术大学)
;
Max Planck Institute for Biological Cybernetics(生物 cybernetics 最大平面研究所)
;
University of Cambridge(剑桥大学)
;
Stanford University(斯坦福大学)
专题命中
推理与问题求解
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
HAMLET: A Hierarchical and Adaptive Multi-Agent Framework for Live Embodied Theatrics
HAMLET:一种分层自适应的多智能体框架用于实时具身戏剧
Shufan Jiang, Sizhou Chen, Chios Chen, Chi Zhang, Xiao-Lei Zhang, Xuelong Li
机构
*
East China University of Science and Technology(东华大学)
;
The University of Sydney(悉尼大学)
;
Institute of Artificial Intelligence (TeleAI)(人工智能研究所)
;
China Telecom(中国电信)
;
Datawhale Org(Datawhale组织)
;
Independent Researcher(独立研究员)
专题命中
推理与问题求解
:large language model(abstract);language model(abstract);分类 cs.AI
Kintsugi: Learning Policies by Repairing Executable Knowledge Bases
Kintsugi: 通过修复可执行知识库学习策略
Teng Cao, Yu Deng, Hikaru Shindo, Quentin Delfosse, Lanxi Wen, Suli Wang, Jannis Blüml, Christopher Tauchmann, Kristian Kersting
机构
*
Artificial Intelligence and Machine Learning Lab, Technical University of Darmstadt, Germany(德累斯顿技术大学人工智能与机器学习实验室)
;
Hessian Center for Artificial Intelligence (hessian.AI), Germany(黑森人工智能中心)
;
Department of Computer Science, Technical University of Darmstadt, Germany(德累斯顿技术大学计算机科学系)
;
Department of Computer Science, Technical University of Munich (TUM), Germany(慕尼黑技术大学计算机科学系)
;
German Research Center for Artificial Intelligence (DFKI), Germany(德国人工智能研究中心)
;
Centre for Cognitive Science, Technical University of Darmstadt, Germany(德累斯顿技术大学认知科学中心)