arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-02-10 至 2026-02-10 共收录 455 信号源:cs.CL, cs.AI, cs.LG

1. 推理与问题求解 79 篇

2602.08804 2026-02-10 cs.AI 89%

Root Cause Analysis Method Based on Large Language Models with Residual Connection Structures

基于大语言模型与残差连接结构的根因分析方法

Liming Zhou, Ailing Liu, Hongwei Liu, Min He, Heng Zhang

专题命中 推理与问题求解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 本文提出一种基于大语言模型与残差连接结构的根因分析方法,通过整合多源遥测数据和建模因果依赖,提升复杂微服务架构中的根因定位能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00380 2026-02-10 cs.CL 89%

Clause-Internal or Clause-External? Testing Turkish Reflexive Binding in Adapted versus Chain of Thought Large Language Models

内部短语还是外部短语?测试土耳其反身代词在适应型与思维链大型语言模型中的绑定

Sercan Karakaş

机构 * University of Chicago(芝加哥大学)

专题命中 推理与问题求解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

AI总结 本研究通过对比两种大型语言模型,探讨土耳其反身代词在不同模型中的绑定行为差异,揭示模型架构和训练数据对指代依赖表示的影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.08085 2026-02-10 physics.soc-ph cs.AI cs.CE 89%

Large language models for spreading dynamics in complex systems

复杂系统中传播动力学的大型语言模型

Shuyu Jiang, Hao Ren, Yichang Gao, Yi-Cheng Zhang, Li Qi, Dayong Xiao, Jie Fan, Rui Tang, Wei Wang

机构 * School of Cyber Science and Engineering, Sichuan University(四川大学计算机科学与工程学院) Key Laboratory of Data Protection and Intelligent Management (Sichuan University), Ministry of Education(数据保护与智能管理重点实验室(四川大学)) Cyber Science Research Institute, Sichuan University(网络科学研究院) Royal Melbourne Institute of Technology University(皇家墨尔本理工大学) Physics Department, University of Fribourg(弗里堡大学物理系) Chongqing Academy of Preventive Medicine(重庆预防医学研究院) School of Public Health and Emergency Management, Chongqing Medical and Pharmaceutical College(重庆医药学院公共卫生与应急管理学院) School of Public Health, Chongqing Medical University(重庆医科大学公共卫生学院)

专题命中 推理与问题求解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 本文探讨了大型语言模型在复杂系统中传播动力学研究中的应用,涵盖数字流行病和生物流行病两个领域,分析了LLMs在流行病建模、检测与预测中的作用及未来研究方向。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.07609 2026-02-10 cs.SE cs.AI 89%

Evaluating Large Language Models for Detecting Architectural Decision Violations

评估大型语言模型在检测架构决策违规方面的表现

Ruoyu Su, Alexander Bakhtin, Noman Ahmad, Matteo Esposito, Valentina Lenarduzzi, Davide Taibi

机构 * University of Oulu, Finland(奥卢大学,芬兰) University of Southern Denmark, Vejle, Denmark(南部丹麦大学,维杰)

专题命中 推理与问题求解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 本研究评估了大型语言模型在检测架构决策违规方面的有效性,发现其在显式决策上表现良好,但在隐式决策上存在局限,需结合专家评估。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.07360 2026-02-10 eess.SY cs.SY 89%

In-Context System Identification for Nonlinear Dynamics Using Large Language Models

基于大语言模型的非线性动力学在境系统识别

Linyu Lin

专题命中 推理与问题求解 :large language model(title,abstract);language model(title,abstract);LLM(abstract)

AI总结 本文提出一种利用大语言模型辅助SINDy的方法,通过在境学习迭代优化方程结构,提升复杂动力学系统的符号恢复精度与可解释性。

Comments 6 pages, 5 figures, submitted to The 10th IEEE Conference on Control Technology and Applications (CCTA) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.11397 2026-02-10 cs.CV 89%

EAGLE: Elevating Geometric Reasoning through LLM-empowered Visual Instruction Tuning

通过LLM赋能的视觉指令微调提升几何推理能力

Zhihao Li, Yao Du, Yang Liu, Yan Zhang, Yufang Liu, Mengdi Zhang, Xunliang Cai, Charles Ling, Boyu Wang

机构 * Department of Computer Science, Western University(计算机科学系,西部大学) Meituan Inc.(美团公司) Department of Automation, Tsinghua University(自动化系,清华大学) School of Computer Science and Technology, East China Normal University(计算机科学与技术学院,东华大学)

专题命中 推理与问题求解 :LLM(title,abstract);instruction tuning(title);large language model(abstract);language model(abstract)

AI总结 EAGLE通过视觉增强框架提升几何推理能力,解决MLLMs在几何理解与视觉感知上的不足。

Comments revised version

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.05239 2026-02-10 cs.RO 89%

Lan-grasp: Using Large Language Models for Semantic Object Grasping and Placement

Lan-grasp:利用大语言模型进行语义物体抓取与放置

Reihaneh Mirjalili, Michael Krawez, Yannik Blei, Simone Silenzi, Florian Walter, Wolfram Burgard

机构 * Department of Computer Science and Artificial Intelligence, University of Technology Nuremberg(计算机科学与人工智能系,图恩技术大学) Department of Engineering ``Enzo Ferrari'' (DIEF), University of Modena and Reggio Emilia(恩佐·费尔拉蒂工程系(DIEF),摩德纳和雷吉奥艾米利亚大学) Technical University of Munich(慕尼黑技术大学)

专题命中 推理与问题求解 :large language model(title,abstract);language model(title,abstract);foundation model(abstract)

AI总结 Lan-grasp利用大语言模型实现更精确的语义抓取与放置,通过结合多种模型提升抓取效果和安全性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.07833 2026-02-10 cs.CV cs.AI cs.CL 88%

SPD-Faith Bench: Diagnosing and Improving Faithfulness in Chain-of-Thought for Multimodal Large Language Models

SPD-Faith Bench: 诊断和改进多模态大语言模型中的推理忠实性

Weijiang Lv, Yaoxuan Feng, Xiaobo Xia, Jiayu Wang, Yan Jing, Wenchao Chen, Bo Chen

机构 * Xidian University(西安电子科技大学) National University of Singapore(新加坡国立大学) Xi’an Jiaotong University(西安交通大学)

专题命中 推理与问题求解 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 SPD-Faith Bench通过细粒度图像差异推理诊断并改进多模态大语言模型中的推理忠实性,揭示了感知盲区和感知-推理脱节的系统性失败模式,并提出SAGE框架提升视觉路由和推理与感知的一致性。

Comments 53 pages, 42 figures, 14 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.08213 2026-02-10 cs.LG cs.AI cs.CL q-bio.QM 88%

DrugR: Optimizing Molecular Drugs through LLM-based Explicit Reasoning

DrugR: 通过基于大语言模型的显式推理优化分子药物

Haoran Liu, Zheni Zeng, Yukun Yan, Yuxuan Chen, Yunduo Xiao

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);pretraining(abstract)

AI总结 DrugR通过基于大语言模型的显式推理方法,优化分子药物的ADMET性质,提升药理效果并提供可解释的优化依据。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.07187 2026-02-10 cs.AI 88%

PreFlect: From Retrospective to Prospective Reflection in Large Language Model Agents

PreFlect:从回顾性到前瞻性反思在大语言模型代理中的应用

Hanyu Wang, Yuanpu Cao, Lu Lin, Jinghui Chen

机构 * College of Information Sciences and Technology, The Pennsylvania State University, State College, PA, USA(信息科学与技术学院,宾夕法尼亚州立大学,州立学院,PA,美国)

专题命中 推理与问题求解 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 PreFlect通过前瞻性反思机制在执行前改进代理计划,提升复杂任务的性能,优于现有反思方法和更复杂的架构。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.22443 2026-02-10 cs.CL 88%

Accounting Reasoning in Large Language Models: Concepts, Evaluation, and Empirical Analysis

在大型语言模型中进行会计推理:概念、评估与实证分析

Jie Zhou, Xin Chen, Jie Zhang, Zhe Li

机构 * School of Computer Engineering, Jiangsu Ocean University(计算机工程学院,江苏海洋大学)

专题命中 推理与问题求解 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 本文研究了会计推理的概念、评估方法及实证分析,发现GPT-4在会计推理任务中表现最佳,但现有LLMs仍需优化以满足实际应用需求。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07820 2026-02-10 cs.AI cs.CL 87%

Certainty-Guided Reasoning in Large Language Models: A Dynamic Thinking Budget Approach

在大语言模型中的确定性引导推理:一种动态思维预算方法

João Paulo Nogueira, Wentao Sun, Alonso Silva, Laith Zumot

机构 * Nokia Bell Labs, France(诺基亚贝尔实验室,法国)

专题命中 推理与问题求解 :language model(title,abstract);large language model(title);分类 cs.CL、cs.AI

AI总结 CGR通过动态调整推理预算,在保持准确性的同时减少token使用,通过确定性评估实现高效的推理过程。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.15586 2026-02-10 cs.CL 85%

Bolmo: Byteifying the Next Generation of Language Models

Bolmo:字节化下一代语言模型

Benjamin Minixhofer, Tyler Murray, Tomasz Limisiewicz, Anna Korhonen, Luke Zettlemoyer, Noah A. Smith, Edoardo M. Ponti, Luca Soldaini, Valentin Hofmann

机构 * Allen Institute for AI(艾伦人工智能研究所) University of Cambridge(剑桥大学) University of Washington(华盛顿大学) University of Edinburgh(爱丁堡大学)

专题命中 推理与问题求解 :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL

AI总结 Bolmo通过两阶段转换将现有子词模型转换为字节级模型,实现性能和细粒度理解的平衡。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21868 2026-02-10 cs.HC cs.CL 85%

What Makes LLM Agent Simulations Useful for Policy Practice? An Iterative Design Study in Emergency Preparedness

是什么使大语言模型代理模拟对政策实践有用?一项关于应急准备的迭代设计研究

Yuxuan Li, Sauvik Das, Hirokazu Shirado

机构 * School of Computer Science, Carnegie Mellon University(卡内基梅隆大学计算机科学学院)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文通过迭代设计研究,探讨了LLM代理模拟在应急准备中的应用,发现通过可验证场景建立信任、获取隐性知识以及共同进化模拟与政策实施是提升其对政策实践有用性的关键。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.21239 2026-02-10 cs.AI 85%

TIDE: Tuning-Integrated Dynamic Evolution for LLM-Based Automated Heuristic Design

TIDE:基于大语言模型的自动启发式设计调谐集成动态进化

Chentong Chen, Mengyuan Zhong, Ye Fan, Jialong Shi, Jianyong Sun

机构 * School of Mathematics and Statistics, Xi’an Jiaotong University, Xi’an, China(西安交通大学数学与统计学院) School of Electronics and Information, Northwest Polytechnical University, Xi’an, China(西北工业大学电子与信息学院)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 TIDE通过集成调谐的动态进化框架,解决大语言模型在自动化启发式设计中结构与参数优化耦合的问题,提升搜索效率和解的质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.02019 2026-02-10 cs.CL 85%

ChatCFD: An LLM-Driven Agent for End-to-End CFD Automation with Structured Knowledge and Reasoning

ChatCFD: 一种基于大语言模型的端到端CFD自动化代理

E Fan, Kang Hu, Zhuowen Wu, Jiangyang Ge, Jiawei Miao, Yuzhi Zhang, He Sun, Weizong Wang, Tianhan Zhang

机构 * Department of Mechanics and Aerospace Engineering, Southern University of Science and Technology(南方科技大学机械与航空航天工程系) School of Astronautics, Beihang University(北航航天学院) DP Technology, Beijing(北京DP技术) College of Future Technology, Peking University(北京大学未来技术学院) National Biomedical Imaging Center, Peking University(北京大学国家生物医学成像中心) State Key Laboratory of High-Efficiency Reusable Aerospace Transportation Technology(高性能可重复使用航天运输技术国家重点实验室) Key Laboratory of Spacecraft Design Optimization and Dynamic Simulation Technology, Ministry of Education(航天器设计优化与动态仿真技术重点实验室)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 ChatCFD通过结构化知识和推理,实现端到端CFD自动化,显著提升执行成功率和物理保真度,具备高灵活性和模块化设计。

Comments 19 pages, 8 figures

Journal ref Adv. Intell. Discov. 2025, 2500174

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.08421 2026-02-10 cs.RO 85%

Decentralized Intent-Based Multi-Robot Task Planner with LLM Oracles on Hyperledger Fabric

基于Hyperledger Fabric的去中心化意图多机器人任务规划器与LLM预言机

Farhad Keramat, Salma Salimi, Tomi Westerlund

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract)

AI总结 本文提出基于Hyperledger Fabric的去中心化多机器人任务规划器,结合LLM预言机与新聚合方法,实现意图分解与多机器人协调。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.08384 2026-02-10 cs.CR 85%

Towards Real-World Industrial-Scale Verification: LLM-Driven Theorem Proving on seL4

迈向现实工业级验证:基于LLM的定理证明在seL4上的应用

Jianyu Zhang, Fuyuan Zhang, Jiayi Lu, Jilin Hu, Xiaoyi Yin, Long Zhang, Feng Yang, Yongwang Zhao

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract)

AI总结 本文提出AutoReal,一种基于LLM的工业级定理证明方法,通过轻量级本地部署和改进的证明训练,实现了在seL4验证项目中51.67%的证明成功率,并在其他安全项目中达到53.88%的证明成功率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21880 2026-02-10 cs.CL cs.AI cs.LG 85%

No Prompt Left Behind: Exploiting Zero-Variance Prompts in LLM Reinforcement Learning via Entropy-Guided Advantage Shaping

无提示被忽视:通过熵引导的优势塑造利用零方差提示进行LLM强化学习

Thanh-Long V. Le, Myeongho Jeon, Kim Vu, Viet Lai, Eunho Yang

专题命中 推理与问题求解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 通过熵引导的优势塑造,RL-ZVP利用零方差提示提升LLM推理能力,显著优于GRPO及其他基线方法。

Comments ICLR 2026 camera-ready version

Journal ref The Fourteenth International Conference on Learning Representations (ICLR 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.07371 2026-02-10 cs.DB 85%

DeepPrep: An LLM-Powered Agentic System for Autonomous Data Preparation

DeepPrep: 一种基于大语言模型的自主数据准备代理系统

Meihao Fan, Ju Fan, Yuxin Zhang, Shaolei Zhang, Xiaoyong Du, Jie Song, Peng Li, Fuxin Jiang, Tieying Zhang, Jianjun Chen

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract)

AI总结 DeepPrep通过基于树状代理推理和渐进式训练框架,实现高效且准确的数据准备,相比闭源模型具有更低的成本和更广的适用性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.07594 2026-02-10 cs.CL cs.AI 84%

Learning to Self-Verify Makes Language Models Better Reasoners

学习自我验证使语言模型更擅长推理

Yuxin Chen, Yu Wang, Yi Zhang, Ziang Ye, Zhengzhou Cai, Yaorui Shi, Qi Gu, Hui Su, Xunliang Cai, Xiang Wang, An Zhang, Tat-Seng Chua

机构 * National University of Singapore(新加坡国立大学) University of Science and Technology of China(中国科学技术大学) Beijing University of Posts and Telecommunications(北京邮电大学)

专题命中 推理与问题求解 :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI

AI总结 通过自我验证学习提升语言模型的推理能力,有效提高生成和验证性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.09000 2026-02-10 cs.AI 83%

iGRPO: Self-Feedback-Driven LLM Reasoning

iGRPO:基于自我反馈的LLM推理

Ali Hatamizadeh, Shrimai Prabhumoye, Igor Gitman, Ximing Lu, Seungju Han, Wei Ping, Yejin Choi, Jan Kautz

机构 * NVIDIA(英伟达)

专题命中 推理与问题求解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 iGRPO通过动态自我条件和两阶段优化,提升LLM在数学推理任务中的表现,实现更准确和一致的解决方案。

Comments Tech report

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.08693 2026-02-10 cs.LG 83%

Reasoning aligns language models to human cognition

推理使语言模型对齐人类认知

Gonçalo Guiomar, Elia Torre, Pehuen Moure, Victoria Shavina, Mario Giulianelli, Shih-Chii Liu, Valerio Mante

机构 * ETH AI Center, Zürich, Switzerland ETH Zürich, Switzerland Institute for Neuroinformatics, University of Zürich \& ETH Zürich, Switzerland University College London, London, United Kingdom

专题命中 推理与问题求解 :language model(title,abstract);large language model(abstract);分类 cs.LG

AI总结 本研究通过主动概率推理任务揭示语言模型与人类在不确定性决策中的差异,发现延长推理能显著提升推理性能,但对信息获取影响有限。

Comments 38 pages, 4 main figures, multiple appendix figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.07499 2026-02-10 cs.CL 83%

Let's Simplify Step by Step: Guiding LLM Towards Multilingual Unsupervised Proficiency-Controlled Sentence Simplification

逐步简化:引导LLM实现多语言无监督能力控制的句子简化

Jingshen Zhang, Xin Ying Qiu, Lifang Lu, Zhuhua Huang, Yutao Hu, Yuechang Wu, JunYu Lu

机构 * Department of Computer Science, School of Information Science and Technology, Guangdong University of Foreign Studies(计算机科学系,信息科学与技术学院,广东外语外贸大学) College of Intelligence and Computing, Tianjin University(智能与计算学院,天津大学) Lionrock AI Lab, China Merchants Research Institute of Advanced Technology(Lionrock AI实验室,中国远洋科技研究院)

专题命中 推理与问题求解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出一种多语言无监督句子简化框架,通过分步骤简化提升控制能力,但发现保持语义忠实仍具挑战性。

Comments Accepted to EACL 2026 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23936 2026-02-10 cs.CL 83%

Do Language Models Update their Forecasts with New Information?

语言模型是否会随着新信息更新其预测?

Zhangdie Yuan, Zifeng Ding, Andreas Vlachos

机构 * Department of Computer Science and Technology, University of Cambridge, Cambridge, United Kingdom(计算机科学与技术系,剑桥大学,剑桥,英国)

专题命中 推理与问题求解 :language model(title,abstract);large language model(abstract);分类 cs.CL

AI总结 研究发现语言模型在接收到新信息时更新预测不一致且保守,需改进机制以更准确整合外部知识。

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.01476 2026-02-10 cs.AI cs.CL cs.LG 82%

Tree Search for Language Model Agents

语言模型代理的树搜索

Jing Yu Koh, Stephen McAleer, Daniel Fried, Ruslan Salakhutdinov

机构 * Carnegie Mellon University(卡内基梅隆大学)

专题命中 推理与问题求解 :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出一种树搜索算法,用于提升语言模型代理在现实网页任务中的表现,实验显示其在成功率上显著优于基线方法。

Comments 13 pages. Models and code available at https://jykoh.com/search-agents

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.05818 2026-02-10 cs.AI cs.DB 81%

TKG-Thinker: Towards Dynamic Reasoning over Temporal Knowledge Graphs via Agentic Reinforcement Learning

TKG-Thinker: 向通过代理强化学习实现动态推理的时序知识图谱迈进

Zihao Jiang, Miao Peng, Zhenyan Shan, Wenjie Xu, Ben Liu, Gong Chen, Ziqi Gao, Min Peng

机构 * School of Computer Science, Wuhan University(武汉大学计算机学院) The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) School of Artificial Intelligence, Wuhan University(武汉大学人工智能学院) Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院)

专题命中 推理与问题求解 :large language model(abstract);language model(abstract);SFT(abstract);prompting(abstract)

AI总结 TKG-Thinker通过代理强化学习实现动态推理,提升时序知识图谱问答的性能和泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03994 2026-02-10 cs.LG cs.AI 81%

Bypassing the Rationale: Causal Auditing of Implicit Reasoning in Language Models

绕过理由:语言模型中隐式推理的因果审计

Anish Sathyanarayanan, Aditya Nagarsekar, Aarush Rathore

机构 * Birla Institute of Technology and Science, Pilani, K. K. Birla Goa Campus(比拉理工学院和科学学院,比里拉班加尔学院)

专题命中 推理与问题求解 :language model(title);prompting(abstract);分类 cs.AI、cs.LG

AI总结 本文通过因果审计揭示语言模型中隐式推理的可信度差异,发现CoT的因果影响在模型和任务间存在显著变化,需通过逐层审计来验证。

Comments Under Review at ICLR, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.15375 2026-02-10 cs.CL eess.AS 79%

STITCH: Simultaneous Thinking and Talking with Chunked Reasoning for Spoken Language Models

STITCH:基于分块推理的 spoken language models 的同时思考与说话

Cheng-Han Chiang, Xiaofei Wang, Linjie Li, Chung-Ching Lin, Kevin Lin, Shujie Liu, Zhendong Wang, Zhengyuan Yang, Hung-yi Lee, Lijuan Wang

机构 * National Taiwan University(国立台湾大学) Microsoft(微软)

专题命中 推理与问题求解 :language model(title,abstract);分类 cs.CL

AI总结 STITCH通过分块推理实现语音语言模型的同时思考与说话,提升推理效率与响应速度。

Comments ICLR 2026 camera-ready version. Project page: https://d223302.github.io/STITCH/

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.08948 2026-02-10 cs.AI cs.CL 79%

CoRefine: Confidence-Guided Self-Refinement for Adaptive Test-Time Compute

CoRefine: 基于置信度的自修正以适应测试时计算

Chen Jin, Ryutaro Tanno, Tom Diethe, Philip Teare

专题命中 推理与问题求解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 CoRefine通过置信度引导的自修正方法,在减少计算开销的同时提升推理准确性,适用于需要适应性测试时计算的场景。

详情

展开后加载摘要…

URL PDF HTML 收藏