Distributional Biases in Post-Training: A Markovian Analysis of Reasoning Trajectories
后训练中的分布偏差:推理轨迹的马尔可夫分析
Dake Bu, Wei Huang, Andi Han, Atsushi Nitanda, Bo Xue, Qingfu Zhang, Hau-San Wong, Taiji Suzuki
机构
*
City University of Hong Kong(香港城市大学)
;
Center for Advanced Intelligence Project, RIKEN(RIKEN高级智能研究中心)
;
The Institute of Statistical Mathematics(统计数学研究所)
;
University of Sydney(悉尼大学)
;
CFAR and IHPC, Agency for Science, Technology and Research (A*STAR)(A*STAR的CFAR和IHPC)
;
Nanyang Technological University(南洋理工大学)
;
The University of Tokyo(东京大学)
How Language Models Fail: Token-Level Signatures of Committed and Persistent Reasoning Failures
语言模型如何失败:承诺性和持续性推理错误的令牌级特征
Tanvi Thoria, Kiana Jafari, Marc R. Schlichting, Mykel J. Kochenderfer
机构
*
Department of Computer Science, Stanford University(计算机科学系,斯坦福大学)
;
Department of Aeronautics and Astronautics, Stanford University(航空航天工程系,斯坦福大学)
COMAP: Co-Evolving World Models and Agent Policies for LLM Agents
COMAP:面向LLM智能体的世界模型与智能体策略协同进化
Youwei Liu, Jian Wang, Hanlin Wang, Wenjie Li
机构
*
Central South University(中南大学)
;
College of Computer Science, Sichuan University(四川大学计算机学院)
;
Department of Computing, The Hong Kong Polytechnic University(香港理工大学计算机系)
SpatiaLab: Can Vision-Language Models Perform Spatial Reasoning in the Wild?
SpatiaLab: 视觉-语言模型能否在真实环境中进行空间推理?
Azmine Toushik Wasi, Wahid Faisal, Abdur Rahman, Mahfuz Ahmed Anik, Munem Shahriar, Mohsin Mahmud Topu, Sadia Tasnim Meem, Rahatun Nesa Priti, Sabrina Afroz Mitu, Md. Iqramul Hoque, Shahriyar Zaman Ridoy, Mohammed Eunus Ali, Majd Hawasly, Mohammad Raza, Md Rizwan Parvez
机构
*
Computational Intelligence and Operations Laboratory(计算智能与运筹实验室)
;
Shahjalal University of Science and Technology(沙赫jalal科技大学)
;
BRAC University(BRAC大学)
;
North South University(北南大学)
;
Monash University(墨尔本大学)
;
Qatar Computing Research Institute(卡塔尔计算研究院)
机构
*
Komaba Institute for Science, Graduate School of Arts and Sciences, The University of Tokyo(东京大学艺术科学研究生院Komaba研究所)
;
Nagoya University(名古屋大学)
;
NARA Institute of Science and Technology (NAIST) / Mohamed bin Zayed University of Artificial Intelligence (MBZUAI)(NAIST科学与技术研究所 / 摩洛哥本·泽德人工智能大学)
;
Mohamed bin Zayed University of Artificial Intelligence (MBZUAI) / RIKEN AIP(摩洛哥本·泽德人工智能大学 / RIKEN AIP)
;
Graduate School of Arts and Sciences, The University of Tokyo / RIKEN AIP / Kyoto University(东京大学艺术科学研究生院 / RIKEN AIP / 京都大学)
专题命中
推理与问题求解
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
机构
*
AAII, University of Technology Sydney, New South Wales, Australia(AAII,悉尼大学,新南威尔士州,澳大利亚)
;
Eindhoven University of Technology, Eindhoven, The Netherlands(埃因霍温理工大学,埃因霍温,荷兰)
;
University of Liverpool, Liverpool, United Kingdom(利物浦大学,利物浦,英国)
;
University of New South Wales, New South Wales, Australia(新南威尔士大学,新南威尔士州,澳大利亚)
专题命中
推理与问题求解
:foundation model(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
Exploring Knowledge Conflicts for Faithful LLM Reasoning: Benchmark and Method
探索知识冲突以实现忠实的LLM推理:基准与方法
Tianzhe Zhao, Jiaoyan Chen, Shuxiu Zhang, Haiping Zhu, Qika Lin, Jun Liu
机构
*
School of Computer Science and Technology, Xi'an Jiaotong University(西安交通大学计算机科学与技术学院)
;
Department of Computer Science, The University of Manchester(曼彻斯特大学计算机科学系)
;
Hunan University(湖南大学)
;
National University of Singapore(新加坡国立大学)
专题命中
推理与问题求解
:LLM(title);large language model(abstract);language model(abstract);prompting(abstract)