arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12266 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12266 篇

2307.08234 2023-08-04 eess.AS 88%

Adapting Large Language Model with Speech for Fully Formatted End-to-End Speech Recognition

Shaoshi Ling, Yuxuan Hu, Shuangbei Qian, Guoli Ye, Yao Qian, Yifan Gong, Ed Lin, Michael Zeng

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.10793 2023-07-21 cs.SE 88%

Addressing Compiler Errors: Stack Overflow or Large Language Models?

Patricia Widjojo, Christoph Treude

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.04280 2023-07-11 cs.HC 88%

Shaping the Emerging Norms of Using Large Language Models in Social Computing Research

Hong Shen, Tianshi Li, Toby Jia-Jun Li, Joon Sung Park, Diyi Yang

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.01142 2023-07-04 cs.HC 88%

Prompt Middleware: Mapping Prompts for Large Language Models to UI Affordances

Stephen MacNeil, Andrew Tran, Joanne Kim, Ziheng Huang, Seth Bernstein, Dan Mogil

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.13157 2023-04-27 cs.IR 88%

Generative Relevance Feedback with Large Language Models

Iain Mackie, Shubham Chatterjee, Jeffrey Dalton

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract)

Comments SIGIR 2023 Preprint, 6 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.06597 2023-04-14 cs.HC 88%

"What It Wants Me To Say": Bridging the Abstraction Gap Between End-User Programmers and Code-Generating Large Language Models

Michael Xieyang Liu, Advait Sarkar, Carina Negreanu, Ben Zorn, Jack Williams, Neil Toronto, Andrew D. Gordon

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.04309 2023-04-11 cs.SE 88%

Large Language Models for Business Process Management: Opportunities and Challenges

Maxim Vidgof, Stefan Bachhofner, Jan Mendling

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.02207 2023-04-06 cs.DS cs.CC 88%

Algorithm and Hardness for Dynamic Attention Maintenance in Large Language Models

Jan van den Brand, Zhao Song, Tianyi Zhou

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.15324 2023-03-28 cs.RO 88%

Can Large Language Models design a Robot?

Francesco Stella, Cosimo Della Santina, Josie Hughes

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract)

Comments Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.09185 2023-02-21 cs.CL cs.AI cs.LG 88%

Bounding the Capabilities of Large Language Models in Open Text Generation with Prompt Constraints

Albert Lu, Hongxin Zhang, Yanzhe Zhang, Xuezhi Wang, Diyi Yang

专题命中 其他LLM :large language model(title);language model(title);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 27 pages, 13 figures, 11 tables, to be published in EACL 2023 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.10583 2023-01-03 cs.SE 88%

Automated Repair of Programs from Large Language Models

Zhiyu Fan, Xiang Gao, Martin Mirchev, Abhik Roychoudhury, Shin Hwei Tan

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract)

Comments 12 pages, To appear in ICSE 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.00828 2022-02-03 cs.CL cs.AI cs.LG 88%

Co-training Improves Prompt-based Learning for Large Language Models

Hunter Lang, Monica Agrawal, Yoon Kim, David Sontag

专题命中 其他LLM :large language model(title);language model(title);prompting(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 17 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.02969 2021-12-07 cs.SE cs.PL 88%

Jigsaw: Large Language Models meet Program Synthesis

Naman Jain, Skanda Vaidyanath, Arun Iyer, Nagarajan Natarajan, Suresh Parthasarathy, Sriram Rajamani, Rahul Sharma

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract)

Comments Accepted to ICSE'22

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.21156 2026-08-27 cs.IR cs.AI cs.ET 版本更新 87%

Graph Engineering in the Era of LLM Agents: From Individual Intelligence to System Intelligence

大语言模型智能体时代的图工程:从个体智能到系统智能

Yuyuan Feng, Zhishang Xiang, Chaobin Yang, Qichao Ma, Zerui Chen, Yujing Zhang, Ke Huang, Chuanjie Wu, Zhaoxu Liu, Yili Wang, Xin He, Jiapu Wang, Zijin Hong, Hao Chen, Yuanchen Bei, Kun Wang, Shengyuan Chen, Ningyu Zhang, Enyan Dai, Linhao Luo, Qingyi Pan, Qi Wang, Wenqi Fan, Guangjing Wang, Na Zou, Yangqiu Song, Xin Wang, Zechao Li, Xia Hu, Qing Li, Xiao Huang, Zhihong Zhang, Jinsong Su, Qinggang Zhang, Yi Chang

专题命中 其他LLM :LLM(title,summary_cn);分类 cs.AI

AI总结 该研究针对LLM智能体复杂任务的个体智能局限,提出图工程范式,通过构建动态图结构组织协调异构智能体,为系统智能提供统一基础。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.23941 2026-08-26 cs.AI 新提交 87%

More Rejective, Not More Discriminative: The Unit of Verification in Pre-Execution LLM Oversight

更多弃权(不执行),而非更强判别性:预执行大语言模型监督中的验证单元

Yuchen Han, Cheng Yan, Wuyang Zhang

机构 * University of Science and Technology of China(中国科学技术大学)

专题命中 其他LLM :LLM(title,summary_cn);分类 cs.AI

AI总结 本研究针对预执行LLM监督的验证单元问题,提出双前缀框架,发现一或两个行动的短验证单元信息性最高,更长窗口会使监控器更倾向于弃权而非提升判别性。

Comments 37 pages, 20 figures, 32 tables, including appendices

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.23023 2026-08-26 cs.CL 版本更新 87%

Most of the LLM Routing Gap Is Task Type

大语言模型路由差距主要源于任务类型

Janghoon Lee

机构 * Redrob(雷德罗布)

专题命中 其他LLM :LLM(title,summary_cn);分类 cs.CL

AI总结 本研究发现LLM路由的性能差距主要源于任务类型,提前为每种任务类型分配固定模型的静态策略,成本低于最佳单一模型,可优化大部分问题,而学习型路由器的优化目标问题数量少于运行间波动。

Comments 21 pages, 2 figures, 12 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.20438 2026-08-24 physics.soc-ph cs.AI cs.MA cs.SI 新提交 87%

Peer-Voted LLM-Agent Stress Tests Find Feed-Induced Lexical Convergence but No Reliable Matched-Exposure Advantage for Distributed Sources

同行投票的大语言模型智能体压力测试:发现feed诱导的词汇趋同,但分布式来源无可靠的匹配暴露优势

Rana Muhammad Usman, Dominic Williamson

专题命中 其他LLM :LLM(title,summary_cn);分类 cs.AI

AI总结 本研究通过PV-SST测试床开展实验,发现同行排名feed会诱导LLM智能体群体的词汇趋同,但未证实分布式来源相比单一来源具有可靠的立场改变优势。

Comments 9 pages, 3 figures. Code, frozen protocol, configurations, summary tables, and data are publicly available at https://github.com/ranausmanai/synthetic-social-networks and https://huggingface.co/datasets/ranausmans/synthetic-social-networks

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.11207 2026-08-13 cs.AI 新提交 87%

Dynamic Governance of Multi-LLM Agent Systems for Collaborative Conversational Outcomes

用于实现协作式对话结果的多大型语言模型智能体系统的动态管控

Alexander Liss, Nicholas Desmond, Santiago Gil Gallego

专题命中 其他LLM :LLM(title,summary_cn);分类 cs.AI

AI总结 本文针对多LLM智能体对话崩溃问题,提出体验编排器EO管控层,通过三种机制提升顾问联系率,在6万次模拟中取得显著效果,为多智能体协作提供新方案。

Comments 13 pages, 3 figures, 3 tables. Submitted to AI Engineer World's Fair 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.28646 2026-08-12 cs.HC cs.AI 交叉投稿 87%

"YES! YES! I absolutely love this insight!" Affirmative Narration as Interactional Strategy in Dialogues with LLM Chatbots

“是的!是的!我完全喜欢这个洞见!”:对话式大型语言模型聊天机器人交互策略中的肯定叙事

Hanna-Riikka Roine, Anne Sigrid Refsum, Jill Walker Rettberg

专题命中 其他LLM :LLM(title,summary_cn);分类 cs.AI

AI总结 该研究分析了LLM聊天机器人对话中最大化用户参与度的肯定叙事交互策略,指出其含引导角色、激活masterplots等机制,还揭示了该叙事的风险,提出需培养新型素养。

Comments In review for a special issue of Narrative Inquiry

Journal ref 2026 Narrative Inquiry 23 (2)

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.18414 2026-08-11 cs.CR cs.AI 版本更新 87%

Prompts Don't Protect: Architectural Enforcement via MCP Proxy for LLM Tool Access Control

提示不保护:通过MCP代理实现的架构强制以实现LLM工具访问控制

Rohith Uppala

机构 * Independent Researcher(独立研究员)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出了一种受控的MCP代理,通过在工具发现和工具调用两个阶段实施基于属性的访问控制(ABAC),有效阻止了未经授权的工具调用,而提示基于的限制仅能减少11-18个百分点的未授权调用率,证明了架构强制在部署的智能体系统中实现可靠工具访问控制的必要性。

Comments 7 pages, 4 tables, 2 figures. Revised version with 200 adversarial tasks and expanded limitations

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01243 2026-08-05 cs.CL 版本更新 87%

Suffix-Constrained Greedy Search Algorithms for Causal Language Models

后缀约束的贪心搜索算法用于因果语言模型

Ayoub Hammal, Pierre Zweigenbaum, Caio Corro

机构 * Université Paris-Saclay, CNRS, LISN(巴黎萨克雷大学、法国国家科学研究中心、LISN) Université de Rennes, INSA Rennes, CNRS, IRISA(雷恩大学、里昂国立应用科学学院、法国国家科学研究中心、IRISA)

专题命中 其他LLM :language model(title,abstract);LLM(abstract,abstract_cn);large language model(abstract);分类 cs.CL

AI总结 本研究提出后缀约束贪心搜索算法,用于在因果语言模型中确保最终答案的结构化提取,同时不损害性能甚至提升结果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.28116 2026-07-28 cs.CL 版本更新 87%

Mechanism-Driven Monitors for Preemptive Detection of LLM Training Instability

机制驱动的LLM训练不稳定性抢先检测监控器

Ruixuan Huang, Hantao Huang, Yifan Huang, Ansheng You, Zhenxing Zhang, Shuai Wang

机构 * HKUST(香港科技大学) Huawei(华为) Independent Researcher(独立研究者)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 针对大语言模型训练中的数值或超参数故障,提出基于模块功能角色的内部监控器,通过QK双线性分解谱熵和MoE路由器指标,在损失发散前数千步检测到不稳定性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.20933 2026-07-24 cs.SE cs.CL 新提交 87%

Transformer-Assisted LLM-Based Source Code Summarisation: to Enable More Secure Software Development

基于Transformer辅助的基于大语言模型的源代码摘要:助力更安全的软件开发

Jesse Phillips, Tracy Hall, Paul Rayson, Mo El-Haj

机构 * School of Computing and Communications, Lancaster University, UK(兰卡斯特大学计算与通讯学院) College of Engineering & Computer Science, VinUniversity(Vin大学工程与计算机科学学院)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 研究旨在解决软件系统中源代码摘要缺失、不完整或过时的问题,通过结合特定任务Transformer模型和大语言模型,利用Transformer生成的摘要辅助提示工程,使LLMs创建更好的源代码摘要,提升了BLEU - 4等指标。

Comments 10 pages

Journal ref NLPAICS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19967 2026-07-23 physics.soc-ph cs.AI cs.CY 新提交 87%

When Shippers Become Algorithms: Candidate Exposure, Information Design, and the Concentration of LLM-Mediated Freight Markets

当托运人成为算法:候选者曝光、信息设计与大语言模型介导的货运市场集中度

Takahiro Ezaki, Naoto Imura, Katsuhiro Nishinari

机构 * Research Center for Advanced Science and Technology, The University of Tokyo(东京大学先进科学与技术研究中心) Department of Aeronautics and Astronautics, School of Engineering, The University of Tokyo(东京大学工学部航空宇宙学系)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究托运人委托大语言模型代理选择承运人对货运市场的影响及平台设计应对策略,通过基于代理的模拟发现代理趋同、集中度随候选列表数量变化等风险,披露承运人剩余日运力可有效应对,凸显平台信息设计的重要性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19300 2026-07-22 cs.AI cs.GT 新提交 87%

LLM Detection as an Intervention: Downstream Impact under Strategic User Behavior

作为一种干预手段的大语言模型检测:战略用户行为下的下游影响

Meena Jagadeesan, Tatsunori Hashimoto, Jon Kleinberg

机构 * Stanford University(斯坦福大学) University of Pennsylvania(宾夕法尼亚大学) Cornell University(康奈尔大学)

专题命中 其他LLM :LLM(title,summary_cn);分类 cs.AI

AI总结 研究LLM检测对下游指标的影响,开发程式化模型捕捉用户策略行为。发现不完善的检测器会扭曲LLM使用及输出质量,导致用户增加使用量且输出质量可能降低,还呈现“先升后降”检测模式,揭示了检测干预的失效模式。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.17708 2026-07-21 cs.AI 新提交 87%

LaT: LLM-as-Trainer for Multi-Task Vehicle Routing Solvers

LaT:用于多任务车辆路径规划求解器的大语言模型训练器

Yang Wang, Ya-Hui Jia, Wei-Neng Chen, Yi Mei, Wen Song, Zhiguang Cao

机构 * School of Future Technology, South China University of Technology(华南理工大学未来技术学院) School of Computer Science and Engineering, South China University of Technology(华南理工大学计算机科学与工程学院) School of Engineering and Computer Science, Victoria University of Wellington(惠灵顿维多利亚大学工程与计算机科学学院) Institute of Marine Science and Technology, Shandong University(山东大学海洋科学与技术研究院) School of Computing and Information Systems, Singapore Management University(新加坡管理大学计算与信息系统学院)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究针对多任务车辆路径规划求解器中VRP变体优化难度不同及现有方法不足的问题,提出LaT训练范式,用预训练大语言模型作外部训练器,定期分析指标生成指导向量注入编码器层,实验证明该范式能提升求解器质量。

Comments 28 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.03695 2026-07-07 cs.LG 新提交 87%

Social Networks of LLM Agents

大语言模型智能体的社交网络

Kaixuan Liu, Guojun Xiong, Weinan Zhang, Shengpu Tang

机构 * Department of Computer Science(计算机科学系) Emory University(埃默里大学) School of Computer Science(计算机科学学院) Shanghai Jiao Tong University(上海交通大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 研究大语言模型智能体群体中集体信念形成问题,介绍SNLA框架,该框架考虑智能体实际影响力,理论与实证验证窄注意力导致羊群效应,宽注意力在特定条件下恢复群体智慧行为。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.03414 2026-07-07 cs.CL 新提交 87%

The Classics at SemEval-2026 Task 3: Combining Transformer Models and LLM-Generated Annotations for Dimensional Aspect-Based Sentiment Analysis

2026年语义评价任务3中的经典方法:结合变压器模型和大语言模型生成的注释进行维度方面的情感分析

Rafif Alshawi, Amit Raj, Aleksey Kudelya, Alexander Shirnin

机构 * HSE University(俄罗斯高等经济研究大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 研究超越传统分类情感,预测情感“效价”和“唤醒”细粒度实值分数的方法。回归任务用基于变压器的编码器模型加权集成,俄语用大语言模型增强输入,提取任务微调解码器大语言模型进行结构化预测。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01133 2026-07-01 cs.CR cs.LG cs.MA 版本更新 87%

When Embedding-Based Defenses Fail: Rethinking Safety in LLM-Based Multi-Agent Systems

基于嵌入的防御失效:重新思考基于大语言模型的多智能体系统的安全性

Lingxi Zhang, Guangtao Zheng, Hanjie Chen

机构 * Rice University(稻属大学) University of Virginia(弗吉尼亚大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 研究探讨了基于嵌入的防御在多智能体系统中的失效模式,提出利用置信度信号提升系统鲁棒性,通过实验验证了早期干预的重要性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.24163 2026-06-24 cs.CR cs.CL 新提交 87%

CORE-BREW: LLR-Based Soft Decoding for Robust Multi-Bit LLM Watermarking

CORE-BREW:基于LLR的软解码用于鲁棒多比特LLM水印

Joeun Kim, HoEun Kim, Young-Sik Kim

机构 * Department of AI DGIST(人工智能系DGIST)

专题命中 其他LLM :LLM(title,title_cn);分类 cs.CL

AI总结 提出CORE-BREW方法,通过固定命中率校准水印信道并计算逐token对数似然比,实现软解码,在保持语义质量的同时提升多比特水印在编辑攻击下的低FPR区分性和鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏