arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

大模型推理能力

大模型数学、逻辑、规划、多步推理和测试时计算能力。

2025-10-14 至 2025-10-14 共收录 21 信号源:cs.CL, cs.AI, cs.LG

1. 其他推理 21 篇

2510.10103 2025-10-14 cs.CL 89%

Stop When Enough: Adaptive Early-Stopping for Chain-of-Thought Reasoning

Renliang Sun, Wei Cheng, Dawei Li, Haifeng Chen, Wei Wang

机构 * UCLA NEC Labs America Arizona State University(亚利桑那州立大学)

专题命中 其他推理 :reasoning(title,abstract);chain-of-thought(title,abstract);CoT(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.14858 2025-10-14 cs.AI cs.CL 81%

Retrieval is Not Enough: Enhancing RAG Reasoning through Test-Time Critique and Optimization

Jiaqi Wei, Hao Zhou, Xiang Zhang, Di Zhang, Zijie Qiu, Wei Wei, Jinzhe Li, Wanli Ouyang, Siqi Sun

机构 * Zhejiang University(浙江大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) South China University of Technology(华南理工大学) University of British Columbia(不列颠哥伦比亚大学) Fudan University(复旦大学) University of Hong Kong(香港大学)

专题命中 其他推理 :reasoning(title,abstract);分类 cs.CL、cs.AI

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11454 2025-10-14 cs.SD cs.AI 79%

Audio-Maestro: Enhancing Large Audio-Language Models with Tool-Augmented Reasoning

Kuan-Yi Lee, Tsung-En Lin, Hung-Yi Lee

机构 * National Taiwan University(国立台湾大学) ASUS Open Cloud Infrastructure Software Center(ASUS开放云基础设施软件中心)

专题命中 其他推理 :reasoning(title,abstract);分类 cs.AI

Comments 9pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10409 2025-10-14 cs.AI 79%

Trace Length is a Simple Uncertainty Signal in Reasoning Models

Siddartha Devic, Charlotte Peale, Arwen Bradley, Sinead Williamson, Preetum Nakkiran, Aravind Gollakota

机构 * University of Southern California(南加州大学) Stanford University(斯坦福大学) Apple(苹果公司)

专题命中 其他推理 :reasoning(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.03941 2025-10-14 cs.LG cs.AI stat.ME 77%

Discovering and Reasoning of Causality in the Hidden World with Large Language Models

Chenxi Liu, Yongqiang Chen, Tongliang Liu, Mingming Gong, James Cheng, Bo Han, Kun Zhang

机构 * TMLR Group, Hong Kong Baptist University(香港 Baptist 大学 TMLR 组) Mohamed bin Zayed University of Artificial Intelligence(Mohamed bin Zayed 人工智能大学) Department of Philosophy, Carnegie Mellon University(哲学系,卡内基梅隆大学) School of Computer Science, The University of Sydney(悉尼大学计算机科学学院) School of Mathematics and Statistics, The University of Melbourne(墨尔本大学数学与统计学学院) The Chinese University of Hong Kong(香港中文大学)

专题命中 其他推理 :reasoning(title,comments);分类 cs.AI、cs.LG

Comments Extended version of our previous NeurIPS'24 conference paper (arXiv:2402.03941(2)); Chenxi and Yongqiang contributed equally; 78 pages, 44 figures; Project page: https://causalcoat.github.io/discovering-and-reasoning

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.13500 2025-10-14 cs.CL cs.AI cs.LG 75%

Noise Injection Systemically Degrades Large Language Model Safety Guardrails

Prithviraj Singh Shahani, Kaveh Eskandari Miandoab, Matthias Scheutz

机构 * Tufts University(塔夫茨大学)

专题命中 其他推理 :reasoning(abstract);chain-of-thought(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 9 pages,3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.00339 2025-10-14 cs.CL cs.AI 73%

VNJPTranslate: A comprehensive pipeline for Vietnamese-Japanese translation

Hoang Hai Phan, Nguyen Duc Minh Vu, Nam Dang Phuong

专题命中 其他推理 :reasoning(abstract);chain-of-thought(abstract);分类 cs.CL、cs.AI

Comments The paper contains a critical error in Section 3.1, leading to invalid results in Section 3.3. This undermines the main conclusion of the paper. The authors are working on a corrected version, but in the meantime, there is not a quick fix/replacement/update available

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09776 2025-10-14 cs.LG cs.AI stat.ML 73%

Why Do Transformers Fail to Forecast Time Series In-Context?

Yufa Zhou, Yixiao Wang, Surbhi Goel, Anru R. Zhang

机构 * Duke University(杜克大学) University of Pennsylvania(宾夕法尼亚大学)

专题命中 其他推理 :chain-of-thought(abstract);CoT(abstract);分类 cs.AI、cs.LG

Comments Code: https://github.com/MasterZhou1/ICL-Time-Series

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05396 2025-10-14 cs.CL cs.AI cs.MA 62%

Talk Isn't Always Cheap: Understanding Failure Modes in Multi-Agent Debate

Andrea Wynn, Harsh Satija, Gillian Hadfield

机构 * Johns Hopkins University(约翰霍普金斯大学) Vector Institute(向量研究所)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI

Comments ICML MAS Workshop 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11292 2025-10-14 cs.LG cs.AI 62%

LouisKV: Efficient KV Cache Retrieval for Long Input-Output Sequences

Wenbo Wu, Qingyi Si, Xiurui Pan, Ye Wang, Jie Zhang

机构 * Peking University(北京大学) Huawei Technologies Ltd.(华为技术有限公司) Chongqing University of Post and Telecommunications(重庆邮电大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10398 2025-10-14 cs.CL cs.AI 62%

STEAM: A Semantic-Level Knowledge Editing Framework for Large Language Models

Geunyeong Jeong, Juoh Sun, Seonghee Lee, Harksoo Kim

机构 * Konkuk University(韩国康康大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI

Comments Accepted to EMNLP 2025 (Findings)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09893 2025-10-14 cs.CL cs.LG 62%

HIPPD: Brain-Inspired Hierarchical Information Processing for Personality Detection

Guanming Chen, Lingzhi Shen, Xiaohao Cai, Imran Razzak, Shoaib Jameel

机构 * School of Electronics and Computer Science University of Southampton(电子与计算机科学学院 英国南安普顿大学) Department of Computational Biology Mohamed bin Zayed University of Artificial Intelligence(计算生物学系 摩洛哥本·扎耶德人工智能大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13380 2025-10-14 cs.RO cs.AI cs.LG cs.MA cs.SY eess.SY 62%

ASTREA: Introducing Agentic Intelligence for Orbital Thermal Autonomy

Alejandro D. Mousist

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

Comments Accepted for presentation at the European Space Agency's AI Start 2025 Conference (see https://atpi.eventsair.com/ai-star-2025/)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07558 2025-10-14 cs.LG cs.AI 62%

VL Norm: Rethink Loss Aggregation in RLVR

Zhiyuan He, Xufang Luo, Yike Zhang, Yuqing Yang, Lili Qiu

机构 * Microsoft Research(微软研究院) Tsinghua University(清华大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11502 2025-10-14 cs.LG 57%

Learning to Make MISTAKEs: Modeling Incorrect Student Thinking And Key Errors

Alexis Ross, Jacob Andreas

机构 * MIT CSAIL(麻省理工学院计算机科学与人工智能实验室)

专题命中 其他推理 :reasoning(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11143 2025-10-14 cs.AI cs.HC 57%

Spec-Driven AI for Science: The ARIA Framework for Automated and Reproducible Data Analysis

Chuke Chen, Biao Luo, Nan Li, Boxiang Wang, Hang Yang, Jing Guo, Ming Xu

专题命中 其他推理 :reasoning(abstract);分类 cs.AI

Comments 19 pages,5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09815 2025-10-14 cs.CV cs.AI 57%

Towards Understanding Ambiguity Resolution in Multimodal Inference of Meaning

Yufei Wang, Adriana Kovashka, Loretta Fernández, Marc N. Coutanche, Seth Wiener

机构 * University of Pittsburgh(匹兹堡大学) Carnegie Mellon University(卡内基梅隆大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI

Comments Accepted to International Conference on Development and Learning (ICDL) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.10736 2025-10-14 cs.CL 57%

Thinking Inside the Mask: In-Place Prompting in Diffusion LLMs

Xiangqi Jin, Yuxuan Wang, Yifeng Gao, Zichen Wen, Biqing Qi, Dongrui Liu, Linfeng Zhang

专题命中 其他推理 :chain-of-thought(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10955 2025-10-14 cs.IR 50%

HatLLM: Hierarchical Attention Masking for Enhanced Collaborative Modeling in LLM-based Recommendation

Yu Cui, Feng Liu, Jiawei Chen, Canghong Jin, Xingyu Lou, Changwang Zhang, Jun Wang, Yuegang Sun, Can Wang

专题命中 其他推理 :reasoning(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10561 2025-10-14 eess.SP 50%

Large Language Model-Empowered Channel Prediction and Predictive Beamforming for LEO Satellite Communications

Zhixiong Chen, Hyundong Shin, Arumugam Nallanathan, Jonathon Chambers

专题命中 其他推理 :reasoning(abstract)

Comments 14 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.14481 2025-10-14 cs.CV 50%

LSP-ST: Ladder Shape-Biased Side-Tuning for Robust Infrared Small Target Detection

Guoyi Zhang, Siyang Chen, Guangsheng Xu, Han Wang, Donghe Wang, Xiaohu Zhang

机构 * School of Aeronautics and Astronautics, Sun Yat-sen University(航空宇航学院,中山大学) Changchun Institute of Optics, Fine Mechanics and Physics, Chinese Academy of Sciences(长春光学精密机械与物理研究所,中国科学院)

专题命中 其他推理 :reasoning(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏