arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型推理能力

大模型数学、逻辑、规划、多步推理和测试时计算能力。

2025-12-23 至 2025-12-23 共收录 18 信号源:cs.CL, cs.AI, cs.LG

1. 其他推理 18 篇

2509.21559 2025-12-23 cs.CV 90%

X-CoT: Explainable Text-to-Video Retrieval via LLM-based Chain-of-Thought Reasoning

X-CoT:基于LLM链式推理的可解释文本到视频检索

Prasanna Reddy Pulakurthi, Jiamian Wang, Majid Rabbani, Sohail Dianat, Raghuveer Rao, Zhiqiang Tao

机构 * Rochester Institute of Technology(罗切斯特理工学院) DEVCOM Army Research Laboratory(陆军研究实验室)

专题命中 其他推理 :reasoning(title,abstract);CoT(title,abstract);chain-of-thought(title)

AI总结 X-CoT通过基于LLM的链式推理实现文本到视频检索的可解释性,提升检索性能并提供详细的推理依据。

Comments 12 pages, 7 figures. Accepted at EMNLP 2025 (Main Conference)

Journal ref Proc. EMNLP 2025, pages 31172-31183, Suzhou, China, Nov. 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.19651 2025-12-23 cs.CL 85%

Exploring Zero-Shot ACSA with Unified Meaning Representation in Chain-of-Thought Prompting

探索链式思维提示中的统一意义表示进行零样本ACS A

Filippos Ventirozos, Peter Appleby, Matthew Shardlow

机构 * Manchester Metropolitan University(曼彻斯特 Metropolitan 大学) Autotrader Research Group(Autotrader 研究组) Autotrader UK(Autotrader 英国)

专题命中 其他推理 :chain-of-thought(title,abstract);reasoning(abstract);CoT(abstract);分类 cs.CL

AI总结 本文提出了一种基于统一意义表示的链式思维提示方法,用于零样本方面-类别情感分析,探讨了其在不同模型规模上的有效性。

Comments 9 pages, 3 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18174 2025-12-23 cs.LG cs.CR cs.CY 85%

Conscious Data Contribution via Community-Driven Chain-of-Thought Distillation

通过社区驱动的推理链蒸馏实现意识数据贡献

Lena Libon, Meghana Bhange, Rushabh Solanki, Elliot Creager, Ulrich Aïvodji

机构 * ETH Zurich(苏黎世联邦理工学院) ÉTS Montréal, Mila(蒙特利尔ÉTS, Mila) University of Waterloo, Vector Institute(多伦多大学, 向量研究所)

专题命中 其他推理 :chain-of-thought(title,abstract);reasoning(abstract);CoT(abstract);分类 cs.LG

AI总结 本文提出通过社区驱动的推理链蒸馏方法,使低效用社区能生成更符合自身目标的替代模型,提升数据可移植性和用户自主性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.19240 2025-12-23 cs.CL cs.AI 81%

ChemATP: A Training-Free Chemical Reasoning Framework for Large Language Models

ChemATP: 一种无需训练的化学推理框架用于大语言模型

Mingxu Zhang, Dazhong Shen, Qi Zhang, Ying Sun

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Nanjing University of Aeronautics and Astronautics(南京航空航天大学) Shanghai AI Laboratory(上海人工智能实验室)

专题命中 其他推理 :reasoning(title,abstract);分类 cs.CL、cs.AI

AI总结 ChemATP通过构建原子级文本知识库,使冻结的大语言模型能够动态检索和推理化学知识,从而在无需训练的情况下实现高效的化学推理。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18822 2025-12-23 cs.AI cs.CL 81%

AdaCtrl: Towards Adaptive and Controllable Reasoning via Difficulty-Aware Budgeting

AdaCtrl: 通过难度感知预算分配实现适应性与可控性推理

Shijue Huang, Hongru Wang, Wanjun Zhong, Zhaochen Su, Jiazhan Feng, Bowen Cao, Yi R. Fung

机构 * Hong Kong University of Science and Technology(香港科技大学) The Chinese University of Hong Kong(香港中文大学) Peking University(北京大学)

专题命中 其他推理 :reasoning(title,abstract);分类 cs.CL、cs.AI

AI总结 AdaCtrl通过难度感知预算分配实现自适应推理控制,提升模型在不同任务中的效率与效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25510 2025-12-23 cs.LG cs.AR 70%

EEsizer: LLM-Based AI Agent for Sizing of Analog and Mixed Signal Circuit

EEsizer: 基于LLM的模拟和混合信号电路尺寸设计AI代理

Chang Liu, Danial Chitnis

机构 * The University of Edinburgh(爱丁堡大学) The School of Engineering(工程学院) Institute for Integrated Micro and Nano Systems(集成微纳系统研究所)

专题命中 其他推理 :reasoning(abstract);chain-of-thought(abstract);分类 cs.LG

AI总结 EEsizer是一种基于LLM的AI代理,通过整合大语言模型与电路仿真器和自定义数据分析功能,实现了模拟和混合信号电路尺寸设计的自动化,展示了在先进节点上的适应性和鲁棒性。

Journal ref IEEE Transactions on Circuits and Systems I: Regular Papers, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.21728 2025-12-23 cs.CL cs.AI cs.LG 67%

Affective Multimodal Agents with Proactive Knowledge Grounding for Emotionally Aligned Marketing Dialogue

具有主动知识 grounding 的情感多模态代理用于情感对齐的营销对话

Lin Yu, Xiaofei Han, Yifei Kang, Chiung-Yi Tseng, Danyang Zhang, Ziqian Bi, Zhimo Han

机构 * Hunan Police Academy, Department of Criminal Investigation(湖南警察学院犯罪侦查系) Business College, California State University(加州州立大学长滩分校商学院) Northwestern University(西北大学) AI Agent Lab, Vokram Group(Vokram集团AI代理实验室) Beijing University of Technology(北京理工大学) Zheng Zhou University of Light Industry(郑州轻工业大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 AffectMind通过主动知识 grounding 和情绪-意图对齐模型,提升营销对话中的情感一致性与说服效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18950 2025-12-23 cs.LG cs.AI 62%

Learning Hierarchical Procedural Memory for LLM Agents through Bayesian Selection and Contrastive Refinement

通过贝叶斯选择和对比细化学习层次化程序记忆以提升LLM代理

Saman Forouzandeh, Wei Peng, Parham Moradi, Xinghuo Yu, Mahdi Jalili

机构 * School of Engineering, Royal Melbourne Institute of Technology University(皇家墨尔本理工学院工程学院)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

AI总结 MACLA通过贝叶斯选择和对比细化方法,实现LLM代理的高效学习与持续改进,无需参数更新。

Comments Accepted at The 25th International Conference on Autonomous Agents and Multi-Agent Systems (AAMAS 2026). 21 pages including references, with 7 figures and 8 tables. Code is publicly available at the authors GitHub repository: https://github.com/S-Forouzandeh/MACLA-LLM-Agents-AAMAS-Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18440 2025-12-23 cs.CL cs.AI 62%

An Agentic AI Framework for Training General Practitioner Student Skills

一种用于培训全科医学生技能的代理AI框架

Victor De Marez, Jens Van Nooten, Luna De Bruyne, Walter Daelemans

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI

AI总结 本文提出了一种代理AI框架,用于培训全科医学生技能,通过统一病例生成、角色驱动对话和基于标准的评估,提升医学教育中虚拟模拟患者的效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18669 2025-12-23 cs.AI 57%

IntelliCode: A Multi-Agent LLM Tutoring System with Centralized Learner Modeling

IntelliCode:一个具有集中式学习者建模的多智能体LLM辅导系统

Jones David, Shreya Ghosh

机构 * School of Computer Science and Engineering, VIT-AP University(计算机科学与工程学院,VIT-AP大学) School of Electrical and Computer Sciences, Indian Institute of Technology Bhubaneswar(电气与计算机科学学院,印度理工学院布巴内斯瓦尔学院)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI

AI总结 IntelliCode通过集中式学习者建模和多智能体协作,实现透明且可靠的LLM辅导系统,提升学习效率和课程适应性。

Comments Submitted to EACL 2026 System Demonstrations Track. 6 pages (main content), 6 figures, includes appendices

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.16661 2025-12-23 cs.IR cs.AI 57%

Microsoft Academic Graph Information Retrieval for Research Recommendation and Assistance

微软学术图信息检索用于研究推荐与协助

Shikshya Shiwakoti, Samuel Goldsmith, Ujjwal Pandit

机构 * Microsoft(微软)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI

AI总结 本文提出基于注意力的子图检索器,利用图神经网络和大语言模型进行高效信息检索与知识推理。

Comments 5 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.18998 2025-12-23 cs.CL 57%

Mirage of Mastery: Memorization Tricks LLMs into Artificially Inflated Self-Knowledge

幻象的精通:LLMs将记忆技巧误认为是人工提升的自我认知

Sahil Kale

机构 * Pune, India(印度浦那)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL

AI总结 本文研究LLMs将记忆误认为智能的问题,揭示其自我认知的不一致性和缺陷,提出需改进模型自我认知的平衡与一致性以提升AI可信度。

Comments 12 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18489 2025-12-23 cs.AI 57%

Large Language Models as Discounted Bayesian Filters

大语言模型作为折扣贝叶斯滤波器

Jensen Zhang, Jing Yang, Keze Wang

机构 * Sun Yat-sen University(中山大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI

AI总结 本研究提出了一种贝叶斯过滤框架,揭示大语言模型在动态环境中的信念更新机制,并提出提示策略以优化其先验校准。

Comments Under submission

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.17914 2025-12-23 cs.CL cs.MA 57%

Q-KVComm: Efficient Multi-Agent Communication Via Adaptive KV Cache Compression

Q-KVComm: 通过自适应KV缓存压缩实现高效的多智能体通信

Boris Kriuk, Logic Ng

机构 * Department of Computer Science \& Engineering Hong Kong University of Science Department of Physics Hong Kong University of Science

专题命中 其他推理 :reasoning(abstract);分类 cs.CL

AI总结 Q-KVComm通过自适应KV缓存压缩实现多智能体高效通信,提升压缩比与语义保真度。

Comments 7 pages, 4 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.14757 2025-12-23 cs.SE cs.AI 57%

SWE-Synth: Synthesizing Verifiable Bug-Fix Data to Enable Large Language Models in Resolving Real-World Bugs

SWE-Synth:合成可验证的bug修复数据以使大语言模型能够解决现实中的bug

Minh V. T. Pham, Huy N. Phan, Hoang N. Phan, Cuong Le Chi, Tien N. Nguyen, Nghi D. Q. Bui

机构 * FPT Software AI Center, Viet Nam(越南FPT软件AI中心) Nanyang Technological University, Singapore(新加坡南洋理工大学) University of Texas at Dallas, US(美国德克萨斯大学达拉斯分校)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI

AI总结 SWE-Synth通过合成可验证的bug修复数据集,提升大语言模型在解决现实bug中的性能。

Comments Work in progress

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18910 2025-12-23 cs.CV 50%

Delta-LLaVA: Base-then-Specialize Alignment for Token-Efficient Vision-Language Models

Delta-LLaVA: 基于令牌效率的视觉-语言模型对齐方法

Mohamad Zamini, Diksha Shukla

机构 * University of Wyoming(怀俄明大学)

专题命中 其他推理 :reasoning(abstract)

AI总结 Delta-LLaVA通过低秩Delta投影和轻量级Transformer块实现视觉-语言模型的高效对齐,提升推理速度和训练效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18671 2025-12-23 cs.CV 50%

SmartSight: Mitigating Hallucination in Video-LLMs Without Compromising Video Understanding via Temporal Attention Collapse

SmartSight: 通过时间注意力崩溃缓解视频大语言模型中的幻觉而不影响视频理解

Yiming Sun, Mi Zhang, Feifei Li, Geng Hong, Min Yang

专题命中 其他推理 :reasoning(abstract)

AI总结 SmartSight通过时间注意力崩溃技术,在不牺牲视频理解能力的前提下,有效降低视频大语言模型的幻觉问题。

Comments AAAI26 accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18496 2025-12-23 cs.CV 50%

Adaptive-VoCo: Complexity-Aware Visual Token Compression for Vision-Language Models

Adaptive-VoCo: 用于视觉-语言模型的复杂度感知视觉标记压缩

Xiaoyang Guo, Keze Wang

机构 * Sun Yat-sen University(中山大学)

专题命中 其他推理 :reasoning(abstract)

AI总结 Adaptive-VoCo通过动态压缩视觉标记提升视觉-语言模型的效率与鲁棒性

Comments Under submission

详情

展开后加载摘要…

URL PDF HTML 收藏