arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-12-09 至 2025-12-09 共收录 23 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 23 篇

2505.21298 2025-12-09 cs.MA cs.AI cs.LG 90%

Large Language Models Miss the Multi-Agent Mark

大语言模型错失多智能体标记

Emanuele La Malfa, Gabriele La Malfa, Samuele Marro, Jie M. Zhang, Elizabeth Black, Michael Luck, Philip Torr, Michael Wooldridge

机构 * Department of Computer Science, University of Oxford(牛津大学计算机科学系) Department of Informatics, King’s College London(伦敦国王学院信息学院) Department of Engineering, University of Oxford(牛津大学工程系) University of Sussex(苏塞克斯大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI、cs.LG

AI总结 本文指出大语言模型多智能体系统在理论与实践间的差距,强调需整合MAS核心概念以避免误解和错失机会。

Comments NeurIPS 2025 - position track -

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06256 2025-12-09 cs.CL cs.AI 89%

Convergence of Outputs When Two Large Language Models Interact in a Multi-Agentic Setup

当两个大语言模型在多智能体设置中交互时输出的收敛性

Aniruddha Maiti, Satya Nimmagadda, Kartha Veerya Jammuladinne, Niladri Sengupta, Ananya Jana

机构 * West Virginia State University(西弗吉尼亚州立大学) Marshall University(马歇尔大学) Fractal Analytics Inc.(Fractal Analytics公司)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI;LLM(comments)

AI总结 研究了两个大语言模型在多智能体环境中交互时输出的收敛现象,发现对话初期连贯但后期趋于重复,导致相似输出循环。

Comments accepted to LLM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07288 2025-12-09 cs.CL 88%

Investigating Training and Generalization in Faithful Self-Explanations of Large Language Models

探究大语言模型忠实自解释的训练与泛化

Tomoki Doi, Masaru Isonuma, Hitomi Yanaka

机构 * The University of Tokyo(东京大学) Riken(理化学研究所) Tohoku University(东北大学) NII LLMC(日本信息处理学会大语言模型委员会)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 本研究通过训练提升大语言模型的自解释忠实性,并验证其在不同任务和风格中的泛化能力。

Comments To appear in the Proceedings of the Asia-Pacific Chapter of the Association for Computational Linguistics: Student Research Workshop (AACL-SRW 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.00781 2025-12-09 cs.AI 88%

CoP: Agentic Red-teaming for Large Language Models using Composition of Principles

CoP: 用于大型语言模型的代理式红队测试:原理组合

Chen Xiong, Pin-Yu Chen, Tsung-Yi Ho

机构 * The Chinese University of Hong Kong Sha Tin, Hong Kong(香港中文大学(深圳)) IBM Research New York, USA(IBM纽约研究院)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 CoP通过原理组合框架实现LLMs红队测试自动化,发现新型jailbreak提示并显著提升攻击成功率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06991 2025-12-09 cs.CL cs.AI 88%

Prompting-in-a-Series: Psychology-Informed Contents and Embeddings for Personality Recognition With Decoder-Only Models

Prompting-in-a-Series: 基于心理启发内容与嵌入的个性识别解码器-only模型

Jing Jie Tan, Ban-Hoe Kwan, Danny Wee-Kiat Ng, Yan-Chai Hum, Anissa Mokraoui, Shih-Yu Lo

机构 * Laboratoire de traitement et transport de l'information, Université Sorbonne Paris Nord, France(信息处理与传输实验室,索邦巴黎北大学,法国) Institute of Communication Studies, National Yang Ming Chiao Tung University, Taiwan(传播学研究所,国立阳明交通大学,台湾)

专题命中 其他LLM :prompting(title,abstract);LLM(abstract);large language model(abstract);language model(abstract)

AI总结 PICEPR通过心理启发内容与嵌入技术,提升解码器-only模型在个性识别中的性能,实现5-15%的提升。

Comments 16 pages

Journal ref IEEE Transactions on Computational Social Systems, pages 1-15, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07797 2025-12-09 cs.CY 85%

LLM Use for Mental Health: Crowdsourcing Users' Sentiment-based Perspectives and Values from Social Discussions

LLM用于心理健康:通过社交媒体讨论众包用户基于情感的视角和价值观

Lingyao Li, Xiaoshan Huang, Renkai Ma, Ben Zefeng Zhang, Haolun Wu, Fan Yang, Chen Chen

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract)

AI总结 本文研究了LLM聊天机器人在不同心理健康状况下的使用情况,通过众包用户数据揭示了情感、视角和价值观的差异,并提出基于价值敏感设计的LLM优化方案。

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.09596 2025-12-09 cs.LG 84%

Mastering AI: Big Data, Deep Learning, and the Evolution of Large Language Models -- AutoML from Basics to State-of-the-Art Techniques

掌握AI:大数据、深度学习与大型语言模型的演变——从基础到最新技术的AutoML

Pohsun Feng, Ziqian Bi, Yizhu Wen, Benji Peng, Junyu Liu, Caitlyn Heqi Yin, Tianyang Wang, Keyu Chen, Sen Zhang, Ming Li, Jiawei Xu, Ming Liu, Xuanhe Pan, Jinlang Wang, Xinyuan Song, Qian Niu

机构 * National Taiwan Normal University(台湾师范大学) Indiana University(印第安纳大学) University of Hawaii(夏威夷大学) AppCubic Kyoto University(京都大学) University of Wisconsin-Madison(威斯康星大学麦迪逊分校) Georgia Institute of Technology(佐治亚理工学院) Rutgers University(罗格斯大学) Purdue University(普渡大学) Emory University(埃默里大学)

专题命中 其他LLM :large language model(title);language model(title);分类 cs.LG

AI总结 本文系统介绍了AutoML的基础知识、主流工具及最新技术,旨在为AI和机器学习领域提供全面的指导与研究参考。

Comments This book contains 169 pages and 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07312 2025-12-09 cs.AR cs.AI cs.DC 83%

DCO: Dynamic Cache Orchestration for LLM Accelerators through Predictive Management

DCO: 通过预测管理实现LLM加速器的动态缓存编排

Zhongchun Zhou, Chengtao Lai, Yuhang Gu, Wei Zhang

机构 * Department of Electronic and Computer Engineering, The Hong Kong University of Science and Technology(电子与计算机工程系,香港科学与技术大学) School of Electronic Science and Engineering, Southeast University(电子科学与工程学院,东南大学)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 DCO通过预测管理实现LLM加速器的动态缓存编排,利用数据流信息优化缓存替换和旁路决策,提升性能至1.8倍,面积仅0.064mm²。

Comments \c{opyright} 2025 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06910 2025-12-09 cs.HC 82%

Robots with Attitudes: Influence of LLM-Driven Robot Personalities on Motivation and Performance

具备态度的机器人:LLM驱动的机器人个性对动机和性能的影响

Dennis Becker, Kyra Ahrens, Connor Gäde, Erik Strahl, Stefan Wermter

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract)

AI总结 本研究探讨了LLM驱动的机器人顺从性对合作任务表现的影响,发现顺从性可提升受欢迎程度和任务表现。

Journal ref Proceedings of the 13th International Conference on Human-Agent Interaction, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02195 2025-12-09 cs.CL cs.AI cs.LG cs.MA 82%

A Knowledge-Based Language Model: Deducing Grammatical Knowledge in a Multi-Agent Language Acquisition Simulation

基于知识的语言模型:多智能体语言习得模拟中的语法知识推导

David Ph. Shakouri, Crit Cremers, Niels O. Schiller

机构 * Leiden University Centre for Linguistics (LUCL), Leiden University, the Netherlands(莱顿大学语言学中心(LUCL)、莱顿大学、荷兰) Leiden Institute for Brain and Cognition (LIBC), Leiden University, the Netherlands(莱顿脑与认知研究所(LIBC)、莱顿大学、荷兰) City University of Hong Kong (CityU), Hong Kong(香港城市大学(CityU)、香港)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 MODOMA系统通过多智能体模拟实现了基于知识的语言模型,展示了儿童智能体在不同示例数据下习得语法类别的能力。

Comments 23 pages, 7 figures, 11 tables. Related work: arXiv:2503.18702. This is the peer-reviewed publisher's version, downloadable from: https://www.clinjournal.org/clinj/article/view/193

Journal ref Computational Linguistics in the Netherlands Journal, 14, 167-189 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03929 2025-12-09 cs.AI 81%

MOTIF: Multi-strategy Optimization via Turn-based Interactive Framework

MOTIF: 通过回合式交互框架进行多策略优化

Nguyen Viet Tuan Kiet, Dao Van Tung, Tran Cong Dao, Huynh Thi Thanh Binh

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 MOTIF通过回合式交互框架实现多策略优化,提升求解器设计的自动化水平。

Comments Accepted as an oral presentation at AAAI 2026. Code available at: https://github.com/HaiAu2501/MOTIF

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.05630 2025-12-09 cs.CR cs.AI cs.LG 81%

How Not to Detect Prompt Injections with an LLM

如何不通过LLM检测提示注入

Sarthak Choudhary, Divyam Anshumaan, Nils Palumbo, Somesh Jha

机构 * University of Wisconsin-Madison(威斯康星大学麦迪逊分校)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI、cs.LG

AI总结 本文揭示了KAD方案的结构性漏洞,并提出DataFlip攻击方法,能够有效绕过KAD防御,实现低检测率和高恶意行为诱导率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07234 2025-12-09 cs.CV cs.AI 79%

Dropout Prompt Learning: Towards Robust and Adaptive Vision-Language Models

Dropout Prompt Learning: 向稳健和适应性强的视觉-语言模型迈进

Biao Chen, Lin Zuo, Mengmeng Jing, Kunbin He, Yuchen Wang

机构 * Biao Chen, Lin Zuo, Mengmeng Jing, Kunbin He, Yuchen Wang(作者)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

AI总结 Dropout Prompt Learning通过在视觉-语言模型中应用Dropout技术,提升模型的鲁棒性和适应性,实验表明其在低样本学习、长尾分类和分布外泛化等任务中表现优异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07022 2025-12-09 cs.SE cs.AI cs.IR 77%

Reformulate, Retrieve, Localize: Agents for Repository-Level Bug Localization

重新表述、检索、局部化:用于仓库级缺陷定位的代理

Genevieve Caumartin, Glaucia Melo

机构 * Concordia University(康科迪亚大学) Toronto Metropolitan University(多伦多 Metropolitan 大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出利用LLM代理通过查询重新表述和摘要改进仓库级缺陷定位,实现更高效的文件级定位性能。

Comments Accepted at BoatSE 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06836 2025-12-09 cs.SE cs.AI cs.PL 77%

Leveraging LLMs to support co-evolution between definitions and instances of textual DSLs

利用LLMs支持文本DSL定义与实例的共进化

Weixing Zhang, Regina Hebig, Daniel Strüber

机构 * Chalmers University of Technology(查尔姆斯理工大学) University of Gothenburg(哥德堡大学) Radboud University(拉德博德大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本研究利用大型语言模型支持文本DSL定义与实例的共进化,通过实验验证了LLM在迁移文本实例方面的有效性及可扩展性挑战。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07306 2025-12-09 stat.ML cs.AI cs.LG 73%

Exact Synthetic Populations for Scalable Societal and Market Modeling

可扩展的社会与市场建模中的精确合成人口

Thierry Petit, Arnault Pachot

机构 * institutetext: Emotia, France(Emotia机构) STATION F

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本研究提出了一种基于约束编程的合成人口生成方法,实现高精度人口统计控制,适用于社会与市场建模场景。

Comments Submitted for peer review on December 7, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06042 2025-12-09 cs.SE cs.AI 70%

Auto-SPT: Automating Semantic Preserving Transformations for Code

Auto-SPT:用于代码的自动化语义保持变换

Ashish Hooda, Mihai Christodorescu, Chuangang Ren, Aaron Wilson, Kassem Fawaz, Somesh Jha

机构 * Google(谷歌) Google, U. Wisconsin–Madison(谷歌,威斯康星大学麦迪逊分校) U. Wisconsin–Madison, Google(威斯康星大学麦迪逊分校,谷歌)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 Auto-SPT通过生成多样化的语义保持变换来提升代码克隆检测模型的鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07818 2025-12-09 cs.LG cs.AI stat.ML 62%

Provable Long-Range Benefits of Next-Token Prediction

可证明的长距离收益的下一个令牌预测

Xinyuan Cao, Santosh S. Vempala

机构 * Georgia Tech(佐治亚理工学院)

专题命中 其他LLM :language model(abstract);分类 cs.AI、cs.LG

AI总结 该研究证明了下一个令牌预测在学习长距离结构上的强大能力,并提供了模型大小的多项式界来解释实际中观察到的长距离连贯性。

Comments 66 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07360 2025-12-09 cs.CV cs.AI 57%

Structure-Aware Feature Rectification with Region Adjacency Graphs for Training-Free Open-Vocabulary Semantic Segmentation

基于区域邻接图的结构感知特征校正用于无训练开放词汇语义分割

Qiming Huang, Hao Ai, Jianbo Jiao

机构 * The MIx Group, School of Computer Science University of Birmingham(米克集团,计算机科学学院,伯明翰大学)

专题命中 其他LLM :language model(abstract);分类 cs.AI

AI总结 本文提出基于区域邻接图的结构感知特征校正方法,通过增强局部辨别能力来提升开放词汇语义分割的性能。

Comments Accepted to WACV2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02002 2025-12-09 cs.LG cs.CE 57%

Generative Large-Scale Pre-trained Models for Automated Ad Bidding Optimization

生成式大规模预训练模型用于自动化广告竞价优化

Yu Lei, Jiayang Zhao, Yilei Zhao, Zhaoqi Zhang, Linyou Cai, Qianlong Xie, Xingxing Wang

机构 * Beijing University of Posts and Telecommunications(北京邮电大学) Nanyang Technological University(南洋理工大学)

专题命中 其他LLM :foundation model(abstract);分类 cs.LG

AI总结 GRAD通过生成式模型结合专家混合模块和因果变压器,提升广告竞价效率与收益,实现更高 ROI 和 GMV。

Comments KDD 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07599 2025-12-09 cs.CV 50%

Online Segment Any 3D Thing as Instance Tracking

在线分割三维物体作为实例跟踪

Hanshi Wang, Zijian Cai, Jin Gao, Yiwei Zhang, Weiming Hu, Ke Wang, Zhipeng Zhang

机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), CASIA(多模态人工智能系统国家重点实验室(MAIS),中国科学院自动化所) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) AutoLab, School of Artificial Intelligence, Shanghai Jiao Tong University(上海交通大学人工智能学院AutoLab) Anyverse Intelligence Beijing Key Laboratory of Super Intelligent Security of Multi-Modal Information(北京超智能多模态信息安全重点实验室) School of Information Science and Technology, ShanghaiTech University(上海科技大学信息科学与技术学院)

专题命中 其他LLM :foundation model(abstract)

AI总结 将在线3D分割重新定义为实例跟踪问题,通过时间信息传播和空间一致性学习提升具身智能体对环境的理解能力。

Comments NeurIPS 2025, Code is at https://github.com/AutoLab-SAI-SJTU/AutoSeg3D

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07381 2025-12-09 cs.CV 50%

Tessellation GS: Neural Mesh Gaussians for Robust Monocular Reconstruction of Dynamic Objects

Tessellation GS:基于神经网格高斯的鲁棒单目动态物体重建

Shuohan Tao, Boyao Zhou, Hanzhang Tu, Yuwang Wang, Yebin Liu

机构 * University of Cambridge(剑桥大学) Tsinghua University(清华大学)

专题命中 其他LLM :foundation model(abstract)

AI总结 Tessellation GS通过基于网格面的神经高斯方法实现单目动态物体的鲁棒重建,显著提升了稀疏视角和动态场景下的重建性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.16376 2025-12-09 cond-mat.str-el 50%

Stability of algebraic spin liquids coupled to quantum phonons

代数自旋液体与量子声子耦合的稳定性

Francesco Ferrari, Josef Willsher, Urban F. P. Seifert, Roser Valentí, Johannes Knolle

专题命中 其他LLM :prompting(abstract)

AI总结 研究通过变分蒙特卡洛方法探讨了三角晶格上J1-J2反铁磁Heisenberg模型在自旋-声子耦合下的稳定性,发现低温下从U(1) DSL到价键有序的转变,并预测了实现稳定DSL基态的参数范围。

Journal ref Physical Review Research 7, L042053 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏