Selective Forgetting for Large Reasoning Models
选择性遗忘用于大推理模型
机构 * Iowa State University(爱荷华州立大学)
专题命中 其他推理 :reasoning(title,abstract);CoT(abstract);分类 cs.AI
AI总结 本文提出一种选择性遗忘框架,通过多语言模型分析推理轨迹,识别并替换敏感内容,以保护一般推理能力,同时解决大推理模型的知识泄露问题。
AI 大模型
大模型数学、逻辑、规划、多步推理和测试时计算能力。
选择性遗忘用于大推理模型
机构 * Iowa State University(爱荷华州立大学)
专题命中 其他推理 :reasoning(title,abstract);CoT(abstract);分类 cs.AI
AI总结 本文提出一种选择性遗忘框架,通过多语言模型分析推理轨迹,识别并替换敏感内容,以保护一般推理能力,同时解决大推理模型的知识泄露问题。
基于令牌级适应性潜在链式思维的预训练
机构 * LUMIA Lab(LUMIA实验室) ; School of Artificial Intelligence(人工智能学院) ; Shanghai Jiao Tong University(上海交通大学) ; Shanghai Al Laboratory(上海Al实验室) ; Sun Yat-sen University(中山大学) ; Shanghai Innovation Institute(上海创新研究院)
专题命中 其他推理 :chain-of-thought(title,abstract);CoT(abstract);分类 cs.CL
AI总结 本研究提出了一种基于令牌级适应性潜在链式思维的预训练方法,通过内部化潜在链式思维来提高模型性能,减少计算成本,并在语言建模和下游任务中取得显著改进。
Comments 15pages
当思考反噬:对推理引发的不一致的机理洞察
专题命中 其他推理 :reasoning(title,abstract);CoT(abstract);分类 cs.CL
AI总结 本文揭示了推理能力增强导致的不一致现象,通过神经元层面的分析,解释了推理与安全之间的纠缠如何引发灾难性遗忘。
Comments ICLR 2026
解构大音频-语言模型中的推理能力以实现模糊情绪预测
机构 * University of Auckland, New Zealand(奥克兰大学) ; The University of Melbourne, Australia(墨尔本大学) ; University of Birmingham, United Kingdom(伯明翰大学)
专题命中 其他推理 :reasoning(title,abstract);chain-of-thought(abstract);分类 cs.AI
AI总结 本文提出了一种新的方法,通过将模糊情绪识别转化为分布推理问题,提升大音频-语言模型在模糊情绪预测中的推理能力。
Comments The paper was submitted to Interspeech for review
通过检索-过渡头桥接潜在推理与目标语言生成
专题命中 其他推理 :reasoning(title,abstract);chain-of-thought(abstract);分类 cs.CL
AI总结 本研究通过引入检索-过渡头(RTH)揭示多语言模型中负责目标语言生成的关键注意力机制,并通过实验验证其在链式推理中的重要性。
Comments In the paper, there are still many statements that are unclear and lack sufficient justification. Since it is difficult for us to estimate how much time would be required to properly revise and correct these issues, we would like to request a withdrawal of the paper in this moment. Thank you!
ImgCoT:将长链推理压缩为紧凑的视觉标记以实现大语言模型的高效推理
机构 * National University of Defense Technology(国防科技大学)
专题命中 其他推理 :reasoning(title,abstract);CoT(abstract);分类 cs.AI
AI总结 ImgCoT通过将长链推理压缩为视觉标记,提升大语言模型的推理效率,同时保留全局结构和细节信息。
通过多模态推理实现对视觉语言模型的逃逸攻击
机构 * Novi High School, MI, USA(诺维高中) ; Michigan State University, MI, USA(密歇根州立大学)
专题命中 其他推理 :reasoning(title);chain-of-thought(abstract);CoT(abstract);分类 cs.AI
AI总结 本文提出了一种基于多模态推理的逃逸攻击框架,通过训练后链式推理和自适应噪声机制提高对视觉语言模型的安全过滤器的绕过能力。
通过主动安全推理增强模型对劫持的防御
机构 * School of Computing(计算学院) ; National University of Singapore(新加坡国立大学) ; School of Computer Science and Engineering(计算机科学与工程学院) ; Nanyang Technological University(南洋理工大学) ; Department of Computing(计算系) ; The Hong Kong Polytechnic University(香港理工大学)
专题命中 其他推理 :reasoning(title,abstract);chain-of-thought(abstract);分类 cs.AI
AI总结 本文提出SCoT方法,通过主动安全推理提升模型对劫持的防御能力,有效减少对分布外问题和对抗操纵的易感性。
ReaSeq:通过推理解锁世界知识用于序列建模
机构 * TaoRank Team(TaoRank团队)
专题命中 其他推理 :reasoning(title,abstract);chain-of-thought(abstract);分类 cs.CL
AI总结 ReaSeq通过引入世界知识增强推理,提升推荐系统在物品表示和用户兴趣建模上的性能,实现IPV、CTR、订单和GMV的显著提升。
通过反向学习推理步骤来衡量推理链的忠实性
机构 * Technion - Israel Institute of Technology(技术学院-以色列理工学院) ; University of Utah(犹他大学)
专题命中 其他推理 :reasoning(title,abstract);CoT(abstract);分类 cs.CL
AI总结 通过反向学习推理步骤来评估推理链的参数忠实性,揭示模型推理与预测之间的关系。
Comments Outstanding paper at EMNLP 2025
因果强度与渗漏信念:通过噪声-或因果贝叶斯网络解读LLM推理
机构 * New York University(纽约大学)
专题命中 其他推理 :reasoning(title,abstract);CoT(abstract);分类 cs.AI
AI总结 本文通过噪声-或因果贝叶斯网络评估LLM和人类在因果推理任务中的表现,探讨其因果强度和信念差异。
Journal ref WiML Workshop at NeurIPS 2025
使用语言模型可解释性进行无监督解码编码推理
机构 * Goodfire AI ; Anthropic ; Cambridge-Boston Alignment Initiative(剑桥-波士顿对齐计划)
专题命中 其他推理 :reasoning(title,abstract);chain-of-thought(abstract);分类 cs.AI
AI总结 本研究通过微调模型进行加密推理并利用logit lens分析,验证了可解释性技术在解码隐藏推理过程中的有效性,提出了无监督解码方法以提升对AI系统监督能力。
Comments Modifying author contact information
基于大语言模型推理的注意力提示的半导体时间序列回归增强方法
机构 * Engineering Product Development (EPD), SUTD, Singapore(新加坡南洋理工大学工程产品开发部) ; Institute for Infocomm Research (I$^2$R), A*STAR, Singapore(新加坡信息通信研究所)
专题命中 其他推理 :reasoning(title,abstract);chain-of-thought(abstract);分类 cs.LG
AI总结 TS-Hint通过结合大语言模型推理提供注意力提示,提升半导体时间序列回归在有限数据下的性能。
从幻觉到意图:用于视觉-语言推理的视觉理由学习
机构 * Zhejiang University(浙江大学) ; The University of Hong Kong(香港大学) ; Meituan(美团)
专题命中 其他推理 :reasoning(title,abstract);chain-of-thought(abstract);分类 cs.AI
AI总结 本文提出视觉理由学习(ViRL),通过将视觉动作作为核心推理元素,提升视觉-语言推理模型的透明度和可信度。
Comments 19 pages, 15 figures
机构 * Yuhang Wang, Yanxu Zhu, Jitao Sang(作者)
专题命中 其他推理 :reasoning(title,abstract);chain-of-thought(abstract);分类 cs.AI
机构 * The Alan Turing Institute(艾伦·图灵研究所)
专题命中 其他推理 :reasoning(title,abstract);chain-of-thought(abstract);分类 cs.AI
机构 * Microsoft Research Asia(微软亚洲研究院) ; Shanghai Jiao Tong University(上海交通大学) ; Carnegie Mellon University(卡内基梅隆大学)
专题命中 其他推理 :reasoning(title,abstract);chain-of-thought(abstract);分类 cs.CL
机构 * School of Computer Science and Engineering, Central South University(中南大学计算机科学与工程学院) ; Research Center for Social Computing and Interactive Robotics, Harbin Institute of Technology(哈尔滨工业大学社会计算与交互机器人研究中心) ; Institute of Computing and Intelligence, Harbin Institute of Technology, Shenzhen(哈尔滨工业大学深圳研究院计算与智能研究所) ; Text Computing and Cognitive Intelligence Ministry of Education Engineering Research Center, Guizhou University(贵州大学文字计算与认知智能教育部工程研究中心) ; Chinese University of Hong Kong(香港中文大学) ; Shanghai AI Laboratory(上海人工智能实验室) ; National University of Singapore(新加坡国立大学) ; Peking University(北京大学) ; ByteDance Seed (China)(字节跳动种子(中国))
专题命中 其他推理 :chain-of-thought(title,abstract);reasoning(abstract);分类 cs.CL
Comments Accepted at NeurIPS 2025;
机构 * Squirrel Ai Learning ; The Ohio State University(俄亥俄州立大学) ; City St George’s, University of London(伦敦城市圣乔治学院)
专题命中 其他推理 :reasoning(title,abstract);CoT(abstract);分类 cs.CL
Comments Work in Progress
机构 * Inria Paris(巴黎研究所)
专题命中 其他推理 :reasoning(title,abstract);CoT(abstract);分类 cs.CL
机构 * Fudan University(复旦大学) ; Tsinghua University(清华大学) ; Bytedance(字节跳动)
专题命中 其他推理 :reasoning(title,abstract);chain-of-thought(abstract);分类 cs.AI
专题命中 其他推理 :reasoning(title,abstract);CoT(abstract);分类 cs.AI
机构 * Tencent Hunyuan(腾讯文英)
专题命中 其他推理 :reasoning(title,abstract);CoT(abstract);分类 cs.CL
Comments Accepted by EMNLP 2025
机构 * Fudan University(复旦大学) ; Shanghai Jiao Tong University(上海交通大学) ; SIA-Lab of Tsinghua AIR(清华大学AIR实验室)
专题命中 其他推理 :reasoning(title,abstract);chain-of-thought(abstract);分类 cs.CL
专题命中 其他推理 :reasoning(title,abstract);CoT(abstract);分类 cs.CL
Comments Our code is publicly available at https://github.com/yuyijiong/hard_retrieval_for_llm and the datasets is at https://huggingface.co/datasets/yuyijiong/difficult_retrieval
机构 * First Author Affiliation(第一作者机构)
专题命中 其他推理 :chain-of-thought(title,abstract);CoT(abstract);分类 cs.CL
机构 * Truthful AI ; UC Berkeley(加州大学伯克利分校)
专题命中 其他推理 :reasoning(title,abstract);CoT(abstract);分类 cs.LG
Comments 10 pages, 8 figures
机构 * Charles L. Brown Department of Electrical and Computer Engineering, University of Virginia, USA(弗吉尼亚大学电气与计算机工程系) ; Department of Computer Science, University of Virginia, USA(弗吉尼亚大学计算机科学系)
专题命中 其他推理 :chain-of-thought(title);reasoning(abstract);CoT(abstract);分类 cs.LG
机构 * Carnegie Mellon University(卡内基梅隆大学) ; Sony Group Corporation(索尼集团)
专题命中 其他推理 :chain-of-thought(title,abstract);CoT(abstract);分类 cs.CL
Comments Accepted at INTERSPEECH 2025
机构 * East China Normal University(东华大学) ; Youtu Lab, Tencent(腾讯优图实验室) ; Shanghai Normal University(上海师范大学)
专题命中 其他推理 :reasoning(title,abstract);CoT(abstract);分类 cs.CL
Comments Accepted by ACL2025(Findings)