arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12228 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12228 篇

2607.03238 2026-07-07 cs.SE 新提交 75%

An Empirical Study of Downstream Adaptation for Agent Skills

智能体技能下游适配的实证研究

Xinjian Wu, Jingzhi Gong, Gunel Jahangirova, Zhenpeng Chen, Jie M. Zhang

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 研究大语言模型智能体技能下游适配,分析六个技能库的1126个适配实例,构建含46种模式的分类法,揭示重用悖论,发现适配相互依赖及近五分之一适配引入安全敏感内容,为改进技能设计等提供启示。

Comments 10 pages, 3 figures, 13 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.08616 2026-07-07 cs.SE 75%

Towards Automated Identification of Violation Symptoms of Architecture Erosion

朝向自动化识别架构侵蚀违规症状

Ruiyin Li, Peng Liang, Paris Avgeriou, Yifei Wang

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文研究利用传统机器学习、深度学习和大语言模型自动检测代码审查中的架构违规症状,通过实验发现SVM结合词嵌入表现最佳,大语言模型在不平衡数据集上表现更优,研究为提升架构一致性提供了自动化方法。

Comments 47 pages, 6 images, 17 tables, Manuscript revision submitted to a journal (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.13771 2026-07-07 cs.CV 版本更新 75%

AppAgent: Multimodal Agents as Smartphone Users

AppAgent:作为智能手机用户的多模态智能体

Chi Zhang, Zhao Yang, Jiaxuan Liu, Yanda Li, Yucheng Han, Xin Chen, Zebiao Huang, Bin Fu, Gang Yu

机构 * Tencent(腾讯)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 介绍基于大语言模型的多模态智能体框架,通过简化动作空间操作智能手机应用,无需系统后端访问,经自主探索或观察人类示范学习,能执行复杂任务,测试验证其处理多种高级任务的能力。

Comments Accepted to CHI 2025, project page is https://appagent-official.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.27130 2026-07-03 cs.SE 版本更新 75%

A Large-Scale Comprehensive Measurement of AI-Generated Code in Real-World Repositories

对现实仓库中AI生成代码的大型实证研究

Tianhao Mao, Dongfang Zhao, Haixu Tang, Xiaofeng Wang, Hang Zhang

专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract)

AI总结 本文通过大规模实证研究分析现实仓库中AI生成代码的特性,探讨AI辅助开发与传统人类开发的差异,为理解AI在软件开发中的实际影响提供依据。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.30942 2026-07-01 cs.HC cs.CY 新提交 75%

Anthropomorphism in AI Companion Communities: Age, Gender, and Emotional Correlates

AI伴侣社区中的拟人化:年龄、性别与情感相关性

Afia Mubashir, Boden Moraski, Stephanie Choi, Rose E. Guingrich

专题命中 其他LLM :large language model(abstract);language model(abstract);prompting(abstract)

AI总结 本研究利用Reddit数据,分析AI伴侣社区中用户年龄、性别与拟人化倾向及情感表达的关系,发现成年人和女性更易拟人化,积极情感与拟人化正相关,且这些关系在成年人中更强。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.30547 2026-06-30 cs.CY 75%

Teaching Prompt-Based Programming with LLMs: A 45-Minute Lesson with Guided Practice for End-User Programmers

教授基于提示的编程与LLMs:面向最终用户程序员的45分钟指导实践课程

Keith Tran, Samiha Marwan, Thomas Price

专题命中 其他LLM :large language model(abstract);language model(abstract);prompting(abstract)

AI总结 本研究评估了一项45分钟的提示式编程干预课程,通过指导实践提升工程专业学生向LLMs表达计算目标的能力,实验组在提示自我效能上显著提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.30278 2026-06-30 cs.NI 75%

LLMs and Optical Networks: A Symbiotic Relationship

LLMs与光网络:共生关系

Mëmëdhe Ibrahimi, Qiaolun Zhang, Giovanni S. Sticca, Jiaheng Xiong, Francesco Musumeci, Massimo Tornatore

专题命中 其他LLM :LLM(summary_cn,abstract_cn)

AI总结 探讨大语言模型与光网络之间的共生关系,提出LLM训练需要广域网感知的集合通信库、ZR+可插拔光模块和空芯光纤等关键技术,同时LLM可赋能自主网络管理。

Comments 4 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.28920 2026-06-30 cs.CV 75%

ExACT: Exemplar-Driven Calibrated Refinement for Training-Free Visual Grounding in Remote Sensing Images

ExACT: 基于示例驱动的校准精化用于遥感图像中免训练的视觉定位

Zixiao Zhang, Lingling Li, Pei He, Xu Liu, Licheng Jiao

机构 * Xidian University(西安电子科技大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);prompting(abstract)

AI总结 提出ExACT框架,通过一次性视觉提示机制弥合多模态大语言模型在遥感视觉定位中的模态差距,实现免训练的精确像素级定位。

Comments 11 pages, 8 figures, supplementary material included

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02498 2026-06-30 cs.CL cs.AI cs.LG 75%

Test-Time Detoxification without Training or Learning Anything

无需训练或学习的测试时去毒化

Baturay Saglam, Dionysis Kalogerias

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出一种无需训练或学习的测试时去毒化方法,通过零阶优化调整输入嵌入以减少生成文本的毒性,实现安全高效的文本生成。

Comments ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03142 2026-06-30 cs.RO cs.CV 75%

MM-Nav: Multi-View VLA Model for Robust Visual Navigation via Multi-Expert Learning

MM-Nav:基于多专家学习的多视角VLA模型用于鲁棒视觉导航

Tianyu Xu, Jiawei Chen, Jiazhao Zhang, Wenyao Zhang, Zekun Qi, Minghan Li, Zhizheng Zhang, He Wang

机构 * Peking University(北京大学) Galbot Shanghai Jiao Tong University(上海交通大学) Tsinghua University(清华大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);foundation model(abstract)

AI总结 本文提出MM-Nav模型,通过多专家学习从合成数据中学习多样化的导航能力,展示模型在合成环境中的强泛化能力,并在现实世界实验中验证其有效性。

Comments Project page: https://pku-epic.github.io/MM-Nav-Web/

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.24060 2026-06-24 cs.SE 新提交 75%

Collaborative and AI-Supported Requirements Elicitation: An Empirical Study

协作与AI支持的需求获取:一项实证研究

Manoel Salgado Neto, Alan Araujo, Ronnie de Souza Santos

专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract)

AI总结 通过混合方法控制实验,比较四种需求获取方法,发现结合协作与AI支持的方法产出质量最高,且被认为更清晰易执行。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.20595 2026-06-23 cs.HC 新提交 75%

Hybrid Intelligence in Cartoon Captioning: Evaluating AI as a Creative Writing Partner

卡通字幕中的混合智能:评估AI作为创意写作伙伴

Uğur Önal, Sanem Sariel, Metin Sezgin, Derya Akleman, Ergun Akleman

专题命中 其他LLM :large language model(abstract);language model(abstract);prompting(abstract)

AI总结 研究通过GPT-4o为IEEE计算机杂志卡通生成字幕,评估AI在幽默创作中的表现,发现AI能提供创意但需人类把控,建议作为辅助工具。

Comments 12 pages, 8 Figures, Accepted to AI Magazine

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19610 2026-06-16 cs.SE 版本更新 75%

The Influence of Code Comments on the Perceived Helpfulness of Stack Overflow Posts

代码注释对 Stack Overflow 帖子感知有用性的影响

Kathrin Figl, Maria Kirchner, Sebastian Baltes, Michael Felderer

专题命中 其他LLM :large language model(abstract);language model(abstract);prompting(abstract)

AI总结 通过在线实验(n=91)研究代码注释如何影响 Stack Overflow 答案的感知有用性,发现块注释和行内注释均显著提高有用性,且新手认为块注释更有用,而答案位置和分数影响较小。

Comments 32 pages, 7 figures, 2 tables, accepted in Empirical Software Engineering

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04885 2026-06-15 cs.CL cs.AI cs.LG 版本更新 75%

CuMA: Aligning LLMs with Sparse Cultural Values via Demographic-Aware Mixture of Adapters

CuMA: 通过人口统计感知的适配器混合使大语言模型与稀疏文化价值观对齐

Ao Sun, Xiaoyu Wang, Zhe Tan, Yu Li, Jiachen Zhu, Yuheng Jia, Shu Su

机构 * Southeast University(东南大学) ByteDance Inc.(字节跳动公司) Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications (Southeast University), Ministry of Education, China(新一代人工智能技术及其交叉应用重点实验室(东南大学),中华人民共和国教育部,中国)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 提出CuMA框架,通过人口统计感知路由将冲突梯度分离到专家子空间,解决密集模型在多文化对齐中的均值崩溃问题,在WorldValuesBench等基准上取得最优性能。

Comments ACL 2026 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.18934 2026-06-11 cs.HC cs.MM 版本更新 75%

Whispering Water: Materializing Human-AI Dialogue as Interactive Ripples

低语之水:将人机对话物化为交互式涟漪

Ruipeng Wang, Tawab Safi, Yunge Wen, Christina Cunningham, Hoi Ling Tang, Behnaz Farahi

专题命中 其他LLM :LLM(summary_cn,abstract_cn)

AI总结 通过将语音情感转换为激发频率、语义内容输入多智能体LLM系统,以及用对数间距和Bark尺度映射分解合成语音为谐波分量,在物理水面上实现人机对话的物化。

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04809 2026-06-03 cs.CR 75%

SEEM: Exploiting Black-Box Text Attacks to Manipulate Tool Selection

SEEM:利用黑盒文本攻击操纵工具选择

Liuji Chen, Hao Gao, Jinghao Zhang, Qiang Liu, Shu Wu, Liang Wang

专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract)

AI总结 提出SEEM方法,通过词级和字符级的粗到细扰动,在黑盒条件下增加目标工具被选中的概率,揭示工具选择机制的安全漏洞。

Comments 2026 IEEE International Conference on Acoustics, Speech, and Signal Processing

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.29226 2026-05-29 cs.CR 75%

S3C2 Summit 2025-09: Industry Secure Supply Chain Summit

S3C2 峰会 2025-09:行业安全供应链峰会

Md Atiqur Rahman, Yasemin Acar, Michel Cucker, William Enck, Alexandros Kapravelos, Christian Kastner, Dominik Wermke, Laurie Williams

专题命中 其他LLM :LLM(summary_cn,abstract_cn)

AI总结 本文报告了2025年9月由S3C2举办的安全供应链峰会,通过跨行业从业者讨论脆弱依赖、组件选择、恶意提交、构建基础设施、文化和LLM角色六大主题,总结了关键见解和新兴挑战。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.29140 2026-05-29 cs.CR 75%

S3C2 Summit 2025-07: Government Secure Supply Chain Summit

S3C2 2025-07 峰会:政府安全供应链峰会

Sivana Hamer, Pat Morrison, William Enck, Yasemin Acar, Michel Cukier, Alexandros Kapravelos, Christian Kästner, Dominik Wermke, Laurie Williams

专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract)

AI总结 该报告总结了2025年7月9日由S3C2举办的政府安全供应链峰会,来自6个美国联邦机构的12名参与者讨论了软件供应链安全中的SBOM、合规、恶意提交、构建基础设施、文化和大语言模型等六大主题,旨在促进经验分享、形成新合作并指导未来研究方向。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.27643 2026-05-28 cs.RO physics.optics 75%

Agentic Language-to-Objective Synthesis for Optofluidic Assembly

面向光流组件的智能语言到目标合成

Ivan Saraev, Elena Erben, Weida Liao, Fan Nan, Gerhard Neumann, Eric Lauga, Moritz Kreysing

机构 * Institute of Biological and Chemical Systems, Karlsruhe Institute of Technology, Germany(马克斯·普朗克研究所生物和化学系统研究所,卡尔斯鲁厄技术大学,德国) Department of Applied Mathematics and Theoretical Physics, University of Cambridge, UK(应用数学和理论物理系,剑桥大学,英国) Department of Mathematics, Imperial College London, UK(数学系,伦敦帝国理工学院,英国) Institute of Anthropomatics and Robotics (IAR), Karlsruhe Institute of Technology, Germany(人机学与机器人研究所(IAR),卡尔斯鲁厄技术大学,德国)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 提出Speak-to-Objective模块化智能流水线,利用条件大语言模型将口语或书面指令转换为可微目标函数,实现光流控微粒子组装,并支持用户反馈学习。

Comments 21 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.02453 2026-05-18 cs.LG cs.AI cs.CL 75%

How to Train Your Advisor: Steering Black-Box LLMs with Advisor Models

如何训练你的导师:通过导师模型引导黑盒大语言模型

Parth Asawa, Alan Zhu, Abigail O'Neill, Matei Zaharia, Alexandros G. Dimakis, Joseph E. Gonzalez

机构 * University of California, Berkeley(加州大学伯克利分校) Bespoke Labs(Bespoke实验室)

专题命中 其他LLM :language model(abstract);prompting(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出Advisor Models,通过训练小型开放权重模型生成动态个性化建议,提升黑盒前沿模型性能,实验显示在多个任务中效果显著,且具有良好的迁移性和鲁棒性。

Comments International Conference on Machine Learning (ICML) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.11574 2026-05-13 cs.CL cs.AI cs.LG 75%

Three Regimes of Context-Parametric Conflict: A Predictive Framework and Empirical Validation

冲突情境下的三种参数化冲突模式:预测框架与实证验证

Pruthvinath Jeripity Venkata

机构 * Independent Researcher(独立研究者)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出三种冲突处理模式框架,通过实验证明参数强度和唯一性对模型决策的影响,验证了不同模型在不同任务中的表现差异。

Comments 10 pages, 13 tables, no figures. 9,970 API calls across five frontier models

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.06294 2026-05-08 cs.CL cs.AI cs.LG 75%

Log-Likelihood, Simpson's Paradox, and the Detection of Machine-Generated Text

对数似然、辛普森悖论与机器生成文本的检测

Tom Kempton, Viktor Drobnyi, Maeve Madigan, Stuart Burrell

机构 * Department of Mathematics(数学系) University of Manchester(曼彻斯特大学) Risk and Security AI Lab(风险与安全AI实验室) Visa Inc.(Visa公司)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文通过分析对数似然信号在隐藏空间中的非均匀性,提出基于贝叶斯决策理论的局部校准方法,改进文本检测性能,揭示现有检测器的不足并提供可扩展的解决方案。

Comments 10 pages, 3 figures, 2 tables, 11 appendices

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16453 2026-04-30 cs.SE 75%

Understanding the Challenges and Opportunities of Generative AI Apps: An Empirical Study

理解生成式AI应用的挑战与机遇:一项实证研究

Buthayna AlMulla, Maram Assi, Safwat Hassan

专题命中 其他LLM :LLM(abstract,abstract_cn);prompting(abstract)

AI总结 本研究通过分析171款生成式AI应用的103万条评论,提出SARA框架,揭示用户对AI功能的认知与评价,识别三大机遇与三大挑战,并揭示用户关切随时间变化的趋势。

Comments 46 pages, 13 figures, 13 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.19429 2026-04-22 cs.HC cs.CY 75%

Discerning Authorship in Online Health Communities: Experience, Trust, and Transparency Implications for Moderating AI

在在线健康社区中辨别作者身份:经验、信任与透明度对调节AI的影响

Yefim Shulman, Agnieszka Kitkowska, Mark Warner

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 研究探讨在线健康社区中用户辨别AI生成建议作者身份的能力,发现透明度与信任的重要性,尽管用户难以区分AI与人类生成内容,但健康状况有显著影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.10575 2026-04-14 cs.HC 75%

NexusAI: Enabling Design Space Exploration of Ideas through Cognitive Abstraction and Functional Decomposition

NexusAI: 通过认知抽象和功能分解实现想法的设计空间探索

Anqi Wang, Bingqian Wang, Huiyang Chen, Keqing Jiao, Lei Han, Xin Tong, Pan Hui

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 NexusAI通过认知抽象和功能分解解决LLM生成想法的结构性不透明问题,提升设计空间探索效率和创造性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.10222 2026-04-14 cs.CY 75%

Morally Programmed LLMs Reshape Human Morality

道德编程的大型语言模型重塑人类道德

Pengzhao Lyu, Yeun Joon Kim, Yingyue Luna Luan, Jungmin Choi

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 研究通过道德编程的LLM与人类交互,发现其能系统性地改变人类道德倾向,且影响持续两周,揭示了道德原则嵌入LLM的伦理困境。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.09444 2026-04-13 cs.HC 75%

Confidence Without Competence in AI-Assisted Knowledge Work

人工智能辅助知识工作中缺乏能力的自信

Elena Eleftheriou, George Pallis, Marios Constantinides

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文研究了不同LLM交互设计对深度思考的影响,发现未来自我解释能提高理解和学习效果,而引导提示能带来最大学习收益。

Comments 25 pages, 13 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.09120 2026-04-13 cs.SE 75%

The Role of LLMs in Collaborative Software Design

大型语言模型在协作软件设计中的作用

Victoria Jackson, Yoonha Cha, Rafael Prikladnicki, André van der Hoek

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 研究探讨了LLM在软件设计协作中的影响,发现共享实例促进理解,而并行使用可能导致上下文漂移,专业人员会审查LLM响应以获得设计洞察,但早期锚定可能限制探索。

Comments accepted into the 2nd HumanAISE workshop 2026, to be published in the FSE Companion '26

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05876 2026-04-09 eess.SY cs.SY 75%

Context-Aware Model Predictive Control for Microgrid Energy Management via LLMs

基于LLMs的上下文感知模型预测控制用于微电网能源管理

Ruixiang Wu, Jiahao Ai, Tinko Sebastian Bartels, Tongxin Li

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出InstructMPC框架,利用LLM和可调最后一层映射,将非结构化操作上下文转化为MPC控制器的预测扰动轨迹,通过理论分析和实验验证,证明了在微电网中整合语义信息提升控制效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.24895 2026-04-08 cs.HC 75%

PII Shield: A Browser-Level Overlay for User-Controlled Personal Identifiable Information (PII) Management in AI Interactions

PII Shield:一种浏览器层面的叠加层,用于在AI交互中实现用户控制的个人可识别信息(PII)管理

Max Holschneider, Saetbyeol LeeYouk

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出PII Shield,通过企业级红actions管道,为用户提供直观的免费AI交互隐私保护体验,引入本地实体匿名化和干扰第三方分析的'烟雾'机制,以平衡用户数据使用与隐私保护。

Comments Accepted at the Proceedings of the CHI 2026 Workshop: Ethics at the Front-End

详情

展开后加载摘要…

URL PDF HTML 收藏