arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型推理能力

大模型数学、逻辑、规划、多步推理和测试时计算能力。

2025-12-30 至 2025-12-30 共收录 12 信号源:cs.CL, cs.AI, cs.LG

1. 其他推理 12 篇

2512.21257 2025-12-30 cs.IR cs.CL 83%

ReaSeq: Unleashing World Knowledge via Reasoning for Sequential Modeling

ReaSeq:通过推理解锁世界知识用于序列建模

Jiakai Tang, Chuan Wang, Gaoming Yang, Han Wu, Jiahao Yu, Jian Wu, Jianwu Hu, Junjun Zheng, Longbin Li, Shuwen Xiao, Xiangheng Kong, Yeqiu Yang, Yuning Jiang, Ahjol Nurlanbek, Binbin Cao, Bo Zheng, Fangmei Zhu, Gaoming Zhou, Huimin Yi, Huiping Chu, Jin Huang, Jinzhe Shan, Kenan Cui, Longbin Li, Silu Zhou, Wen Chen, Xia Ming, Xiang Gao, Xin Yao, Xingyu Wen, Yan Zhang, Yiwen Hu, Yulin Wang, Ziheng Bao, Zongyuan Wu

机构 * TaoRank Team(TaoRank团队)

专题命中 其他推理 :reasoning(title,abstract);chain-of-thought(abstract);分类 cs.CL

AI总结 ReaSeq通过引入世界知识增强推理,提升推荐系统在物品表示和用户兴趣建模上的性能,实现IPV、CTR、订单和GMV的显著提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13797 2025-12-30 cs.CL 79%

Breadcrumbs Reasoning: Memory-Efficient Reasoning with Compression Beacons

痕迹推理:通过压缩信标实现的内存高效推理

Giovanni Monea, Yair Feldman, Shankar Padmanabhan, Kianté Brantley, Yoav Artzi

机构 * Cornell University(康奈尔大学) Harvard University(哈佛大学)

专题命中 其他推理 :reasoning(title,abstract);分类 cs.CL

AI总结 本研究提出通过压缩信标实现内存高效推理,利用联合蒸馏和强化学习框架优化缓存压缩,提升大语言模型在长上下文推理中的内存与准确性平衡。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07583 2025-12-30 cs.CL cs.AI 62%

Complementary Learning Approach for Text Classification using Large Language Models

基于大语言模型的文本分类互补学习方法

Navid Asgari, Benjamin M. Cole

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI

AI总结 本文提出了一种基于大语言模型的文本分类互补学习方法,通过人机协作弥补各自弱点,以低成本技术处理评分差异问题。

Comments After further review, we identified substantive issues that materially affect the validity of the manuscript's core results and conclusions. Addressing these would require a fundamental reworking of the analysis and framing. To maintain the integrity of the public record, we request withdrawal of this version

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01956 2025-12-30 cs.AI cs.LG cs.MA 62%

Scaling Clinician-Grade Feature Generation from Clinical Notes with Multi-Agent Language Models

通过多智能体语言模型实现临床笔记中临床级特征生成的扩展

Jiayi Wang, Jacqueline Jil Vallon, Nikhil V. Kotha, Neil Panjwani, Xi Ling, Margaret Redfield, Sushmita Vij, Sandy Srinivas, John Leppert, Mark K. Buyyounouski, Mohsen Bayati

机构 * Department of Management Science and Engineering, Stanford University School of Engineering(管理科学与工程系,斯坦福大学工程学院) Department of Radiation Oncology, Stanford University School of Medicine(放射肿瘤学系,斯坦福大学医学院) Operations, Information and Technology, Stanford University Graduate Business School(运营、信息与技术,斯坦福大学商学院) Graduate Business School Research Hub, Stanford University Graduate Business School(商学院研究中心,斯坦福大学商学院) Department of Medicine (Oncology), Stanford University School of Medicine(医学系(肿瘤学),斯坦福大学医学院) Department of Medicine, Stanford University School of Medicine(医学系,斯坦福大学医学院) Department of Urology, Stanford University School of Medicine(泌尿学系,斯坦福大学医学院) Veterans Affairs Palo Alto Health Care System(退伍军人事务帕洛阿尔托医疗系统) Department of Electrical Engineering, Stanford University School of Engineering(电气工程系,斯坦福大学工程学院)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

AI总结 本研究提出了一种多智能体语言模型系统,通过自动化临床笔记特征生成,实现了与人工方法相当的预测性能,并在不同医疗场景中展示了良好的可扩展性和可解释性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.15759 2025-12-30 cs.CL cs.AI cs.CV 62%

Vision Enhancing LLMs: Empowering Multimodal Knowledge Storage and Sharing in LLMs

视觉增强大语言模型:赋能大语言模型中的多模态知识存储与共享

Yunxin Li, Zhenyu Liu, Baotian Hu, Wei Wang, Yuxin Ding, Xiaochun Cao, Min Zhang

机构 * Research Institute of Computing and Intelligence(计算与智能研究 institute) Harbin Institute of Technology(哈尔滨工业大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI

AI总结 本文提出MKS2方法,通过多模态知识存储与共享增强大语言模型的推理能力,提升其在物理和常识知识场景下的表现。

Comments 21 pages, 7 figures; Accepted by IEEE TIP

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.22508 2025-12-30 cs.LG cs.AI 62%

Predicting LLM Correctness in Prosthodontics Using Metadata and Hallucination Signals

利用元数据和幻觉信号预测牙科修复学中大语言模型的正确性

Lucky Susanto, Anasta Pranawijayana, Cortino Sukotjo, Soni Prasad, Derry Wijaya

机构 * 1 Department of Data Science, Monash University Indonesia, Tangerang, Indonesia 2 Independent Researcher 3 Department of Prosthodontics, University of Pittsburgh, Pittsburgh, Pennsylvania 4 Department of Restorative Sciences, University of North Carolina Adams School of Dentistry, Chapel Hill, North Carolina 5 Department of Computer Science, Boston University, Boston, Massachusetts

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

AI总结 本文研究通过元数据和幻觉信号预测牙科修复学中LLM的正确性,发现元数据方法可提升准确性,但需进一步改进以适应高风险应用。

Comments Accepted as a Short Paper at HEALTHINF2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.23480 2025-12-30 cs.CR cs.AI 57%

Agentic AI for Autonomous Defense in Software Supply Chain Security: Beyond Provenance to Vulnerability Mitigation

面向软件供应链安全的代理AI:超越溯源到漏洞缓解

Toqeer Ali Syed, Mohammad Riyaz Belgaum, Salman Jan, Asadullah Abdullah Khan, Saad Said Alqahtani

机构 * Faculty of Computer and Information System(计算机与信息系统学院) Islamic University of Madinah(麦地那伊斯兰大学) Faculty of Computer Studies(计算机研究学院) Arab Open University-Bahrain(巴林阿拉伯开放大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI

AI总结 本文提出基于代理AI的软件供应链安全框架,结合LLM推理、强化学习和多代理协调,实现主动漏洞缓解,提升检测准确率和响应效率。

Comments Conference paper, accept in ACCA IEEE Bahrain

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.23430 2025-12-30 cs.CL 57%

C2PO: Diagnosing and Disentangling Bias Shortcuts in LLMs

C2PO:诊断和解构大语言模型中的偏见捷径

Xuan Feng, Bo An, Tianlong Gu, Liang Chang, Fengrui Hao, Peipeng Yu, Shuai Zhao

机构 * Jinan University(济南大学) Nanyang Technological University(南洋理工大学) Engineering Research Center of Trustworthy AI (Ministry of Education)(可信人工智能工程研究中心) Guangxi Key Laboratory of Trusted Software(广西可信软件重点实验室)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL

AI总结 C2PO通过因果对比偏好优化框架,解决大语言模型中的偏见问题,同时保持推理能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.18186 2025-12-30 cs.SD cs.CL eess.AS 57%

Steering Language Model to Stable Speech Emotion Recognition via Contextual Perception and Chain of Thought

通过上下文感知和推理链引导语言模型实现稳定的语音情感识别

Zhixian Zhao, Xinfa Zhu, Xinsheng Wang, Shuiyuan Wang, Xuelong Geng, Wenjie Tian, Lei Xie

专题命中 其他推理 :CoT(abstract);分类 cs.CL

AI总结 C$^2$SER通过上下文感知和推理链提升语音情感识别的稳定性和准确性,优于现有模型。

Comments This work has been published in IEEE Transactions on Audio, Speech and Language Processing

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.01386 2025-12-30 cs.CL cs.CR cs.IR 57%

Topic-FlipRAG: Topic-Orientated Adversarial Opinion Manipulation Attacks to Retrieval-Augmented Generation Models

Topic-FlipRAG: 面向主题的对抗性观点操控攻击用于检索增强生成模型

Yuyang Gong, Zhuo Chen, Jiawei Liu, Miaokun Chen, Fengchang Yu, Wei Lu, Xiaofeng Wang, Xiaozhong Liu

机构 * Wuhan University(武汉大学) Nanyang Technological University(南洋理工大学) Worcester Polytechnic Institute(沃思堡理工学院)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL

AI总结 本文提出Topic-FlipRAG,一种针对检索增强生成模型的面向主题对抗性观点操控攻击方法,通过两阶段流程影响模型输出观点,揭示了RAG系统安全防护的迫切需求。

Comments Accepted by USENIX Security 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.23483 2025-12-30 cs.CV 50%

TV-RAG: A Temporal-aware and Semantic Entropy-Weighted Framework for Long Video Retrieval and Understanding

TV-RAG:一种具有时间意识和语义熵权的长视频检索与理解框架

Zongsheng Cao, Yangfan He, Anran Liu, Feng Chen, Zepeng Wang, Jun Xie

机构 * Researcher(研究者)

专题命中 其他推理 :reasoning(abstract)

AI总结 TV-RAG通过时间衰减检索和熵加权关键帧采样,提升长视频检索与理解性能,无需重新训练即可集成至现有模型。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.01681 2025-12-30 physics.flu-dyn 50%

Large Language Model Driven Development of Turbulence Models

基于大语言模型的湍流模型开发

Zhongxin Yang, Yuanwei Bin, Yipeng Shi, Xiang I. A. Yang

专题命中 其他推理 :reasoning(abstract)

AI总结 本文提出利用大语言模型开发湍流模型,通过闭环迭代流程生成可解释且性能更优的近壁湍流模型,解决了不利压力梯度、系统旋转和表面粗糙度等问题。

Journal ref Flow (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏