arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 18857 信号源:cs.CL, cs.AI, cs.LG

1. 推理与问题求解 18857 篇

2511.22176 2025-12-01 cs.CL cs.AI 86%

Focused Chain-of-Thought: Efficient LLM Reasoning via Structured Input Information

聚焦式链式思考:通过结构化输入信息实现高效的LLM推理

Lukas Struppek, Dominik Hintersdorf, Hannah Struppek, Daniel Neider, Kristian Kersting

机构 * German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心) Technical University of Darmstadt(德累斯顿技术大学) University of Kassel(卡塞尔大学) TU Dortmund University(Dortmund技术大学) TU Center for Trustworthy Data Science and Security, University Alliance Ruhr(TU可信数据科学与安全中心,鲁尔大学联盟) Hessian Center for AI (Hessian.AI)(黑森人工智能中心(黑森.AI)) Centre for Cognitive Science, Technical University of Darmstadt(认知科学中心,德累斯顿技术大学)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出F-CoT方法,通过结构化输入减少LLM推理中的token使用和延迟,同时保持推理准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20726 2025-11-27 cs.LG cs.AI 86%

Learning from Risk: LLM-Guided Generation of Safety-Critical Scenarios with Prior Knowledge

从风险学习:利用先验知识的LLM引导的安全关键场景生成

Yuhang Wang, Heye Huang, Zhenhua Xu, Kailai Sun, Baoshen Guo, Jinhua Zhao

机构 * Chinese Academy of Sciences, China(中国科学院) Department of Urban Studies and Planning, Massachusetts Institute of Technology, USA(麻省理工学院城市研究与规划系) School of Vehicle and Mobility, Tsinghua University, China(清华大学车辆与移动系统学院)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出利用LLM和CVAE生成安全关键场景的方法,通过知识驱动优化提升自动驾驶系统在罕见高风险事件下的鲁棒性与可控性。

Comments 24 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18313 2025-11-25 cs.CL cs.DB cs.IR cs.LG 86%

Path-Constrained Retrieval: A Structural Approach to Reliable LLM Agent Reasoning Through Graph-Scoped Semantic Search

路径约束检索:一种通过图域语义搜索实现可靠LLM代理推理的结构方法

Joseph Oladokun

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 路径约束检索通过结合结构图约束与语义搜索,提升LLM代理推理的可靠性和一致性。

Comments 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13753 2025-11-19 cs.LG cs.AI cs.CR 86%

Robustness of LLM-enabled vehicle trajectory prediction under data security threats

Feilong Wang, Fuqiang Liu

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments 20 pages, 2 figures, 11 tables, working paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.08824 2025-11-18 cs.RO cs.AI cs.CL cs.CY 86%

LLM-Driven Robots Risk Enacting Discrimination, Violence, and Unlawful Actions

Andrew Hundt, Rumaisa Azeem, Masoumeh Mansouri, Martim Brandão

机构 * Carnegie Mellon University(卡内基梅隆大学) King’s College London(伦敦国王学院) University of Birmingham(伯明翰大学)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Published in International Journal of Social Robotics (2025). 49 pages (65 with references and appendix), 27 Figures, 8 Tables. Andrew Hundt and Rumaisa Azeem are equal contribution co-first authors. The positions of the two co-first authors were swapped from arxiv version 1 with the written consent of all four authors. The Version of Record is available via DOI: 10.1007/s12369-025-01301-x

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.01908 2025-11-13 cs.CR cs.AI cs.LG 86%

UDora: A Unified Red Teaming Framework against LLM Agents by Dynamically Hijacking Their Own Reasoning

Jiawei Zhang, Shuang Yang, Bo Li

机构 * Department of Computer Science, University of Chicago(芝加哥大学计算机科学系) Meta Department of Computer Science, University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校计算机科学系)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.11903 2025-11-11 cs.CL cs.AI 86%

DiLA: Enhancing LLM Tool Learning with Differential Logic Layer

Yu Zhang, Hui-Ling Zhen, Zehua Pei, Yingzhao Lian, Lihao Yin, Mingxuan Yuan, Bei Yu

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments arXiv admin note: text overlap with arXiv:2305.12295 by other authors

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00993 2025-11-04 cs.AI cs.LG 86%

Aligning LLM agents with human learning and adjustment behavior: a dual agent approach

Tianming Liu, Jirong Yang, Yafeng Yin, Manzi Li, Linghao Wang, Zheng Zhu

机构 * Department of Civil and Environmental Engineering, University of Michigan, Ann Arbor, United States(美国密歇根大学土木与环境工程系) Department of Electrical Engineering and Computer Science, University of Michigan, Ann Arbor, United States(美国密歇根大学电气工程与计算机科学系) College of Civil Engineering and Architecture, Zhejiang University, Hangzhou, China(浙江大学建筑工程学院)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments 32 pages, 6 figures, 7 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.20432 2025-11-04 cs.AI cs.CY cs.GT cs.LG 86%

LLM Strategic Reasoning: Agentic Study through Behavioral Game Theory

Jingru Jia, Zehua Yuan, Junhao Pan, Paul E. McNamara, Deming Chen

机构 * University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

专题命中 推理与问题求解 :LLM(title);large language model(abstract);language model(abstract);prompting(abstract)

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26274 2025-10-31 cs.CR cs.CL cs.LG 86%

PVMark: Enabling Public Verifiability for LLM Watermarking Schemes

Haohua Duan, Liyao Xiang, Xin Zhang

机构 * Shanghai Jiao Tong University(上海交通大学) Ant Group(蚂蚁集团)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments This work has been submitted to the IEEE for possible publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26143 2025-10-31 cs.AI cs.CL 86%

Reasoning Curriculum: Bootstrapping Broad LLM Reasoning from Math

Bo Pang, Deqian Kong, Silvio Savarese, Caiming Xiong, Yingbo Zhou

机构 * Salesforce AI Research(Salesforce AI研究院) University of California, Los Angeles(加州大学洛杉矶分校)

专题命中 推理与问题求解 :LLM(title);large language model(abstract);language model(abstract);pretraining(abstract)

Comments 9 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.25595 2025-10-30 cs.CL cs.AI 86%

Communication and Verification in LLM Agents towards Collaboration under Information Asymmetry

Run Peng, Ziqiao Ma, Amy Pang, Sikai Li, Zhang Xi-Jia, Yingzhuo Yu, Cristian-Paul Bara, Joyce Chai

机构 * University of Michigan(密歇根大学) UNC, Chapel Hill(北卡罗来纳大学教堂山分校) Georgia Tech(佐治亚理工学院) Apple(苹果公司) Robert Bosch SRL(博世股份有限公司) Babeş-Bolyai University(巴贝什-博耶亚大学)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Workshop on Multi-Agent System @ ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08146 2025-10-29 cs.LG cs.AI 86%

Think Just Enough: Sequence-Level Entropy as a Confidence Signal for LLM Reasoning

Aman Sharma, Paras Chopra

机构 * Lossfunk

专题命中 推理与问题求解 :LLM(title);large language model(abstract);language model(abstract);post-training(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02366 2025-10-28 cs.LG cs.CL q-fin.TR 86%

Language Model Guided Reinforcement Learning in Quantitative Trading

Adam Darmanin, Vince Vella

机构 * University of Malta(马耳他大学)

专题命中 推理与问题求解 :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL、cs.LG

Comments 12 pages (4 pages appendix and references) and 6 figures. Accepted for presentation at FLLM 2025, Vienna

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22034 2025-10-28 cs.AI cs.LG 86%

LLM-AR: LLM-powered Automated Reasoning Framework

Rick Chen, Joseph Ternasky, Aaron Ontoyin Yin, Xianling Mu, Fuat Alican, Yigit Ihlamur

机构 * University of Oxford(牛津大学) Vela Research(Vela研究公司)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21739 2025-10-28 cs.RO cs.AI cs.CL cs.SY eess.SY 86%

Next-Generation LLM for UAV: From Natural Language to Autonomous Flight

Liangqi Yuan, Chuhao Deng, Dong-Jun Han, Inseok Hwang, Sabine Brunswicker, Christopher G. Brinton

机构 * School of Electrical and Computer Engineering, Purdue University(电子与计算机工程学院,普渡大学) School of Aeronautics and Astronautics, Purdue University(航空与航天学院,普渡大学) Department of Computer Science and Engineering, Yonsei University(计算机科学与工程系,延世大学) Polytechnic Institute, Purdue University(技术学院,普渡大学)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.19099 2025-10-28 cs.LG cs.AI 86%

What Makes a Good Curriculum? Disentangling the Effects of Data Ordering on LLM Mathematical Reasoning

Yaning Jia, Chunhui Zhang, Xingjian Diao, Xiangchi Yuan, Zhongyu Ouyang, Chiyu Ma, Soroush Vosoughi

机构 * Dartmouth College(达特茅斯学院)

专题命中 推理与问题求解 :LLM(title);large language model(abstract);language model(abstract);post-training(abstract)

Comments 8 pages (main text) + 4 pages (appendix), 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.16024 2025-10-28 cs.CL cs.AI 86%

The Lighthouse of Language: Enhancing LLM Agents via Critique-Guided Improvement

Ruihan Yang, Fanghua Ye, Jian Li, Siyu Yuan, Yikai Zhang, Zhaopeng Tu, Xiaolong Li, Deqing Yang

机构 * Fudan University(复旦大学) Tencent Hunyuan(腾讯文言)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21584 2025-10-27 cs.CL cs.AI cs.CY 86%

Empirical Evidence for Alignment Faking in a Small LLM and Prompt-Based Mitigation Techniques

Jeanice Koorndijk

机构 * Seraphion Technology(塞拉菲昂技术)

专题命中 推理与问题求解 :LLM(title);large language model(abstract);language model(abstract);prompting(abstract)

Comments NeurIPS RegML Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.19988 2025-10-24 cs.CL cs.AI 86%

LLM-Augmented Symbolic NLU System for More Reliable Continuous Causal Statement Interpretation

Xin Lian, Kenneth D. Forbus

机构 * Department of Computer Science, McCormick School of Engineering Applied Science, \ University, Evanston, IL 60208

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 18 pages, 2 figures

Journal ref Proceedings of the Twelfth Annual Conference on Advances in Cognitive Systems, Poster Collection (2025) 66-83

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.15216 2025-10-22 cs.LG cs.CL 86%

Soundness-Aware Level: A Microscopic Signature that Predicts LLM Reasoning Potential

Xuansheng Wu, Xiaoman Pan, Wenlin Yao, Jianshu Chen

机构 * University of Georgia(佐治亚大学)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments Pre-print

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12838 2025-10-22 cs.CL cs.AI 86%

A$^2$FM: An Adaptive Agent Foundation Model for Tool-Aware Hybrid Reasoning

Qianben Chen, Jingyi Cao, Jiayu Zhang, Tianrui Qin, Xiaowan Li, King Zhu, Dingfeng Shi, He Zhu, Minghao Liu, Xiaobo Liang, Xin Gui, Ge Zhang, Jian Yang, Yuchen Eleanor Jiang, Wangchunshu Zhou

专题命中 推理与问题求解 :foundation model(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 12 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17921 2025-10-22 cs.CL cs.AI 86%

CLAWS:Creativity detection for LLM-generated solutions using Attention Window of Sections

Keuntae Kim, Eunhye Jeong, Sehyeon Lee, Seohee Yoon, Yong Suk Choi

机构 * Department of Computer Science(计算机科学系) Department of Artificial Intelligence(人工智能系) Department of Future Mobility(未来移动部门)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.15979 2025-10-21 cs.LG cs.AI 86%

Cog-Rethinker: Hierarchical Metacognitive Reinforcement Learning for LLM Reasoning

Zexu Sun, Yongcheng Zeng, Erxue Min, Heyang Gao, Bokai Ji, Xu Chen

机构 * Baidu Inc.(百度公司) Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学人工智能学院)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments 22 Pages, 8 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.10343 2025-10-21 cs.CL cs.AI 86%

Code Execution as Grounded Supervision for LLM Reasoning

Dongwon Jung, Wenxuan Zhou, Muhao Chen

机构 * University of California, Davis(加州大学戴维斯分校) University of Southern California(南加州大学)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10592 2025-10-14 cs.AI cs.CL 86%

A Layered Intuition -- Method Model with Scope Extension for LLM Reasoning

Hong Su

机构 * School of Computer Science, Chengdu University of Information Technology(计算机科学学院,成都信息科技学院)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14756 2025-10-10 cs.LG cs.AI 86%

LLINBO: Trustworthy LLM-in-the-Loop Bayesian Optimization

Chih-Yu Chang, Milad Azvar, Chinedum Okwudire, Raed Al Kontar

机构 * University of Michigan(密歇根大学)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.04311 2025-10-07 cs.AI cs.LG 86%

On the Importance of Task Complexity in Evaluating LLM-Based Multi-Agent Systems

Bohan Tang, Huidong Liang, Keyue Jiang, Xiaowen Dong

机构 * University of Oxford(牛津大学) University College London(伦敦大学学院)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.02816 2025-10-06 cs.AI cs.CL 86%

NCV: A Node-Wise Consistency Verification Approach for Low-Cost Structured Error Localization in LLM Reasoning

Yulong Zhang, Li Wang, Wei Du, Peilin Li, Yuqin Dai Zhiyuan Zhao, Lingyong Fang, Ziniu Liu, Ru Zhang, Huijia Zhu, Gongshen Liu

机构 * Beijing University of Posts and Telecommunications(北京邮电大学) Shanghai Jiao Tong University(上海交通大学) Ant Group(蚂蚁集团) Tsinghua University(清华大学) Inner Mongolia Research Institute of SJTU(内蒙古大学SJTU研究所)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.02377 2025-10-06 cs.CL cs.LG 86%

Uncertainty-Aware Answer Selection for Improved Reasoning in Multi-LLM Systems

Aakriti Agrawal, Rohith Aralikatti, Anirudh Satheesh, Souradip Chakraborty, Amrit Singh Bedi, Furong Huang

机构 * University of Maryland(马里兰大学) Hilabs University of Central Florida(佛罗里达中央大学) Capital One

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏