arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型推理能力

大模型数学、逻辑、规划、多步推理和测试时计算能力。

2025-12-01 至 2025-12-01 共收录 22 信号源:cs.CL, cs.AI, cs.LG

1. 规划推理 22 篇

2511.22532 2025-12-01 cs.CV cs.AI 90%

CoT4AD: A Vision-Language-Action Model with Explicit Chain-of-Thought Reasoning for Autonomous Driving

CoT4AD: 一种具有显式推理链的视觉-语言-动作模型用于自动驾驶

Zhaohui Wang, Tengbo Yu, Hao Tang

机构 * Peking University(北京大学)

专题命中 规划推理 :reasoning(title,abstract);chain-of-thought(title,abstract);CoT(abstract);planning(abstract)

AI总结 CoT4AD通过引入显式推理链提升自动驾驶中的视觉-语言-动作模型的推理能力,实现更精确的因果推理和决策。

Comments 10 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.23186 2025-12-01 cs.RO cs.AI cs.CV 83%

Obstruction reasoning for robotic grasping

机器人抓取中的障碍物推理

Runyu Jiao, Matteo Bortolon, Francesco Giuliari, Alice Fasoli, Sergio Povoli, Guofeng Mei, Yiming Wang, Fabio Poiesi

机构 * Fondazione Bruno Kessler(布鲁诺·科塞勒基金会) University of Trento(特伦托大学)

专题命中 规划推理 :reasoning(title,abstract);planning(abstract);分类 cs.AI

AI总结 UNOGrasp通过视觉-语言模型提升机器人抓取中障碍物推理能力,结合监督与强化学习微调,实现更高效的路径规划和抓取性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.23136 2025-12-01 cs.CL cs.AI 81%

Multi-chain Graph Refinement and Selection for Reliable Reasoning in Large Language Models

多链图细化与选择用于大语言模型中的可靠推理

Yujiao Yang, Jing Lian, Linhui Li

专题命中 规划推理 :reasoning(title,abstract);分类 cs.CL、cs.AI

AI总结 MGRS通过多链图细化与选择提升大语言模型的推理能力与效率,实现更高准确率和更快速度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.21928 2025-12-01 cs.LG cs.AI 81%

Prompted Policy Search: Reinforcement Learning through Linguistic and Numerical Reasoning in LLMs

提示策略搜索:通过语言和数值推理在大语言模型中进行强化学习

Yifan Zhou, Sachin Grover, Mohamed El Mistiri, Kamalesh Kalirathnam, Pratyush Kerhalkar, Swaroop Mishra, Neelesh Kumar, Sanket Gaurav, Oya Aran, Heni Ben Amor

机构 * Interactive Robotics Lab, Arizona State University(亚利桑那州立大学交互机器人实验室) Research & Development, Procter & Gamble(普罗cter与 gamble 研究与发展)

专题命中 规划推理 :reasoning(title,abstract);分类 cs.AI、cs.LG

AI总结 ProPS通过结合语言和数值推理,在大语言模型中实现高效的强化学习,展示了在多个任务中超越传统算法的性能。

Comments In The Thirty-ninth Annual Conference on Neural Information Processing Systems

Journal ref Advances in Neural Information Processing Systems (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.21706 2025-12-01 cs.CL cs.AI 81%

A General Highly Accurate Online Planning Method Integrating Large Language Models into Nested Rollout Policy Adaptation for Dialogue Tasks

一种整合大语言模型的通用高精度在线规划方法用于对话任务的嵌套回滚策略适应

Hui Wang, Fafa Zhang, Xiaoyu Zhang, Chaoxu Mu

专题命中 规划推理 :planning(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出NRPA-GD方法,利用大语言模型实现无需训练的对话策略规划,通过嵌套回滚策略适应在目标导向对话任务中取得优异表现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.22598 2025-12-01 cs.LG 79%

LLM-Cave: A benchmark and light environment for large language models reasoning and decision-making system

LLM-Cave: 一个用于大型语言模型推理和决策系统的基准和轻量环境

Huanyu Li, Zongyuan Li, Wei Huang, Xian Guo

机构 * College of Artificial Intelligence Nankai University(人工智能学院 南开大学)

专题命中 规划推理 :reasoning(title,abstract);分类 cs.LG

AI总结 LLM-Cave提出了一种轻量环境和基准,用于评估大型语言模型的推理和决策能力,展示了不同模型在复杂任务中的表现及改进方法。

Comments 8 pages, 5 figures, ICICN 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.22181 2025-12-01 cs.CV cs.AI cs.RO 79%

MTR-VP: Towards End-to-End Trajectory Planning through Context-Driven Image Encoding and Multiple Trajectory Prediction

MTR-VP: 通过基于上下文的图像编码和多轨迹预测实现端到端轨迹规划

Maitrayee Keskar, Mohan Trivedi, Ross Greer

机构 * Machine Intelligence, Interaction, and Imagination (Mi 3 ) Laboratory(机器智能、交互与想象实验室) University of California, Merced(加州大学默塞德分校) Laboratory for Intelligent & Safe Automobiles (LISA)(智能与安全汽车实验室) University of California, San Diego(加州大学圣地亚哥分校)

专题命中 规划推理 :planning(title,abstract);分类 cs.AI

AI总结 MTR-VP通过基于上下文的图像编码和多轨迹预测实现端到端轨迹规划,利用交叉注意力提升规划性能。

Comments 8 pages, 3 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.22891 2025-12-01 cs.AI cs.CL cs.LG 75%

ORION: Teaching Language Models to Reason Efficiently in the Language of Thought

ORION:教语言模型以高效的方式在思维语言中推理

Kumar Tanmay, Kriti Aggarwal, Paul Pu Liang, Subhabrata Mukherjee

机构 * Harvard University(哈佛大学) Hippocratic AI(希波克拉底AI) Massachusetts Institute of Technology(麻省理工学院)

专题命中 规划推理 :reasoning(abstract);planning(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 ORION通过SLPO优化实现高效压缩推理,提升推理效率和准确性,同时保持高精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.22123 2025-12-01 math.OC cs.SY eess.SY 71%

Model Predictive Path Planning in Navier-Stokes Flow with POD-Based Reduced-Order Models

在纳维-斯托克斯流中基于POD的降阶模型的模型预测路径规划

Adam Waterman, Martin Guay

专题命中 规划推理 :planning(title)

AI总结 本文提出一种结合POD降阶模型与MPC的路径规划方法,用于在纳维-斯托克斯流中实现高效实时控制。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18192 2025-12-01 cs.CV cs.AI 70%

ARIAL: An Agentic Framework for Document VQA with Precise Answer Localization

ARIAL:一个用于文档视觉问答的代理框架,具有精确答案定位

Ahmad Mohammadshirazi, Pinaki Prasad Guha Neogi, Dheeraj Kulshrestha, Rajiv Ramnath

机构 * Ohio State University(俄亥俄州立大学) Flairsoft(Flairsoft公司)

专题命中 规划推理 :reasoning(abstract);planning(abstract);分类 cs.AI

AI总结 ARIAL通过代理协调专门工具,实现文档VQA的高精度和可解释性,取得最佳性能和可解释性成果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.21882 2025-12-01 cs.LG cs.CV 70%

Closed-Loop Transformers: Autoregressive Modeling as Iterative Latent Equilibrium

闭环变换器:将自回归建模作为迭代潜在均衡

Akbar Anbar Jafari, Gholamreza Anbarjafari

机构 * University of Tartu(塔尔图大学) S Holding OÜ(3S控股公司)

专题命中 规划推理 :reasoning(abstract);planning(abstract);分类 cs.LG

AI总结 本文提出均衡变换器(EqT),通过迭代细化潜在表示实现自洽均衡,解决自回归模型中错误传播问题,提升长序列推理和多步规划能力。

Comments 22 pages, 1 figure, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.21726 2025-12-01 cs.CL cs.AI cs.LG 67%

Goal-Directed Search Outperforms Goal-Agnostic Memory Compression in Long-Context Memory Tasks

目标导向搜索在长上下文记忆任务中优于目标无关的记忆压缩

Yicong Zheng, Kevin L. McKee, Thomas Miconi, Zacharie Bugaud, Mick van Gelderen, Jed McCaleb

机构 * Astera Institute(Astera研究所)

专题命中 规划推理 :reasoning(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 SUMER通过目标导向搜索在长上下文记忆任务中超越传统记忆压缩方法,达到SOTA性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.22904 2025-12-01 cs.CL cs.LG 62%

Language-conditioned world model improves policy generalization by reading environmental descriptions

语言引导的世界模型通过阅读环境描述提升策略泛化能力

Anh Nguyen, Stefan Lee

机构 * Oregon State University(俄勒冈州立大学)

专题命中 规划推理 :planning(abstract);分类 cs.CL、cs.LG

AI总结 本文提出LED-WM,通过语言引导的世界模型提升策略泛化能力,无需依赖假设,有效应对新动态和语言描述的未见过游戏。

Comments NeuRIPS 2025. Workshop: LAW 2025: Bridging Language, Agent, and World Models

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.22441 2025-12-01 cs.CR cs.AI cs.CV cs.LG 62%

GEO-Detective: Unveiling Location Privacy Risks in Images with LLM Agents

GEO-Detective:利用LLM代理揭示图像中的位置隐私风险

Xinyu Zhang, Yixin Wu, Boyang Zhang, Chenhao Lin, Chao Shen, Michael Backes, Yang Zhang

机构 * CISPA Helmholtz Center for Information Security(CISPA赫尔姆霍茨信息安全中心) Xi’an Jiaotong University(西安交通大学)

专题命中 规划推理 :reasoning(abstract);分类 cs.AI、cs.LG

AI总结 GEO-Detective利用LLM代理进行图像地理定位推断,通过自适应策略和专用工具提升定位精度并降低隐私风险。

Comments 15 pages with 7 figures and 12 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.21753 2025-12-01 cs.CL cs.AI 62%

Extracting Disaster Impacts and Impact Related Locations in Social Media Posts Using Large Language Models

利用大型语言模型从社交媒体帖子中提取灾害影响及影响相关位置

Sameeah Noreen Hameed, Surangika Ranathunga, Raj Prasanna, Kristin Stock, Christopher B. Jones

专题命中 规划推理 :planning(abstract);分类 cs.CL、cs.AI

AI总结 本文利用大型语言模型从社交媒体中提取灾害影响及受影响位置,提升灾害响应决策效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.18924 2025-12-01 cs.CL cs.AI 62%

Simulated patient systems powered by large language model-based AI agents offer potential for transforming medical education

基于大语言模型的AI代理的模拟患者系统有潜力改变医学教育

Huizi Yu, Jiayan Zhou, Lingyao Li, Shan Chen, Jack Gallifant, Anye Shi, Xiang Li, Jingxian He, Wenyue Hua, Mingyu Jin, Guang Chen, Yang Zhou, Zhao Li, Trisha Gupte, Ming-Li Chen, Zahra Azizi, Qi Dou, Bryan P. Yan, Yongfeng Zhang, Yanqiu Xing, Themistocles L. Danielle S. Bitterman, Themistocles L. Assimes, Xin Ma, Lin Lu, Lizhou Fan

专题命中 规划推理 :reasoning(abstract);分类 cs.CL、cs.AI

AI总结 基于大语言模型的AI代理构建的模拟患者系统在医学教育中展现出高保真度和教育价值,优于人类模拟患者。

Comments 19 pages, 6 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.22737 2025-12-01 cs.AI cs.HC 57%

Agentic AI Framework for Individuals with Disabilities and Neurodivergence: A Multi-Agent System for Healthy Eating, Daily Routines, and Inclusive Well-Being

具有残疾和神经多样性个体的代理AI框架:一个用于健康饮食、日常习惯和包容性福祉的多代理系统

Salman Jan, Toqeer Ali Syed, Gohar Ali, Ali Akarma, Mohammad Riyaz Belgaum, Ahmad Ali

专题命中 规划推理 :reasoning(abstract);分类 cs.AI

AI总结 本文提出一种多代理系统,通过个性化营养、适应性调度、食品指导和生理监测等代理,为残疾和神经多样性个体提供健康饮食、日常习惯和包容性福祉的AI框架。

Comments Presented at International Conference on Business and Digital Technology, Bahrain, Springer Nature, 27 November 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.05408 2025-12-01 cs.CR cs.AI cs.CY 57%

Frontier AI's Impact on the Cybersecurity Landscape

前沿人工智能对网络安全领域的冲击

Yujin Potter, Wenbo Guo, Zhun Wang, Tianneng Shi, Hongwei Li, Andy Zhang, Patrick Gage Kelley, Kurt Thomas, Dawn Song

机构 * UC Berkeley(加州大学伯克利分校) UC Santa Barbara(加州大学圣芭芭拉分校) Google(谷歌)

专题命中 规划推理 :planning(abstract);分类 cs.AI

AI总结 本文研究了前沿人工智能在网络安全中的影响,指出人工智能在攻击中的能力已超过防御,呼吁构建新基准、开发防御代理等以缓解风险。

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.16450 2025-12-01 cs.IR 50%

360Brew: A Decoder-only Foundation Model for Personalized Ranking and Recommendation

360Brew:一种用于个性化排序和推荐的解码器-only基础模型

Hamed Firooz, Maziar Sanjabi, Adrian Englhardt, Aman Gupta, Ben Levine, Dre Olgiati, Gungor Polatkan, Iuliia Melnychuk, Karthik Ramgopal, Kirill Talanine, Kutta Srinivasan, Luke Simon, Natesh Sivasubramoniapillai, Necip Fazil Ayan, Qingquan Song, Samira Sriram, Souvik Ghosh, Tao Song, Vignesh Kothapalli, Xiaoling Zhai, Ya Xu, Yu Wang, Yun Dai

专题命中 规划推理 :reasoning(abstract)

AI总结 360Brew是一种基于解码器-only架构的大型基础模型,能够解决多个排序和推荐任务,无需特征工程,性能媲美现有系统。

Comments arXiv admin note: This version has been removed by arXiv administrators as the submitter did not have the right to agree to the license at the time of submission

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.22371 2025-12-01 cs.LO 50%

Hyperintensional Intention

超内涵意图

Daniil Khaitovich, Aybüke Özgün

专题命中 规划推理 :reasoning(abstract)

AI总结 本文提出了一种超内涵意图逻辑,通过避免等价封闭来解决意图推理中的有效性问题,并提供了该逻辑的完备公理化系统。

Comments In Proceedings TARK 2025, arXiv:2511.20540

Journal ref EPTCS 437, 2025, pp. 48-64

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19057 2025-12-01 cs.IR 50%

RELATE: Relation Extraction in Biomedical Abstracts with LLMs and Ontology Constraints

RELATE:利用LLMs和本体约束进行生物医学摘要中的关系提取

Olawumi Olasunkanmi, Mathew Satusky, Hong Yi, Chris Bizon, Harlin Lee, Stanley Ahalt

专题命中 规划推理 :reasoning(abstract)

AI总结 RELATE通过结合LLMs和本体约束,提升生物医学摘要中关系提取的标准化和准确性,有效构建标准化知识图谱。

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13914 2025-12-01 physics.med-ph 50%

Development of a defacing algorithm to protect the privacy of head and neck cancer patients in publicly-accessible radiotherapy datasets

头颈癌患者在公共放射治疗数据集中的隐私保护去脸算法开发

Kayla O'Sullivan-Steben, Luc Galarneau, John Kildea

专题命中 规划推理 :planning(abstract)

AI总结 本研究开发了一种新的自动去脸算法,用于在保护头颈癌患者隐私的同时保留关键结构,使公共放射治疗数据集能够用于大数据和人工智能研究。

Comments 22 pages, 6 figures, submitted to Medical Physics

Journal ref Med Phys. 2025; 52:e70160

详情

展开后加载摘要…

URL PDF HTML 收藏