arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12266 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12266 篇

2307.10169 2023-07-20 cs.CL cs.AI cs.LG 89%

Challenges and Applications of Large Language Models

Jean Kaddour, Joshua Harris, Maximilian Mozes, Herbie Bradley, Roberta Raileanu, Robert McHardy

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments 72 pages. v01. Work in progress. Feedback and comments are highly appreciated!

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.00461 2023-07-04 cs.CL cs.AI cs.LG cs.MM cs.SD 89%

Conformer LLMs -- Convolution Augmented Large Language Models

Prateek Verma

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments 6 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.08161 2023-06-19 cs.CL cs.AI cs.HC cs.IR cs.LG 89%

h2oGPT: Democratizing Large Language Models

Arno Candel, Jon McKinney, Philipp Singer, Pascal Pfeiffer, Maximilian Jeblick, Prithvi Prabhu, Jeff Gambera, Mark Landry, Shivam Bansal, Ryan Chesler, Chun Ming Lee, Marcos V. Conde, Pasha Stetsenko, Olivier Grellier, SriSatish Ambati

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments Work in progress by H2O.ai, Inc

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.11384 2023-06-16 cs.SE 89%

Large Language Models are Few-Shot Summarizers: Multi-Intent Comment Generation via In-Context Learning

Mingyang Geng, Shangwen Wang, Dezun Dong, Haotian Wang, Ge Li, Zhi Jin, Xiaoguang Mao, Xiangke Liao

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract)

Comments Accepted by the 46th International Conference on Software Engineering (ICSE 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.04563 2023-06-08 cs.AI cs.CL cs.HC cs.LG 89%

ChatGPT is fun, but it is not funny! Humor is still challenging Large Language Models

Sophie Jentzsch, Kristian Kersting

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.01773 2023-06-06 cs.CY 89%

Voluminous yet Vacuous? Semantic Capital in an Age of Large Language Models

Luca Nannini

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.06530 2023-05-12 cs.CL cs.AI cs.LG 89%

How Good are Commercial Large Language Models on African Languages?

Jessica Ojo, Kelechi Ogueji

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments Presented at the AfricanNLP Workshop at ICLR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.06030 2023-04-19 cs.CY 89%

The Role of Large Language Models in the Recognition of Territorial Sovereignty: An Analysis of the Construction of Legitimacy

Francisco Castillo-Eslava, Carlos Mougan, Alejandro Romero-Reche, Steffen Staab

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract)

Comments European Workshop of Algorithmic Fairness'23

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.10089 2023-03-20 cs.RO 89%

LP-SLAM: Language-Perceptive RGB-D SLAM system based on Large Language Model

Weiyi Zhang, Yushi Guo, Liting Niu, Peijun Li, Chun Zhang, Zeyu Wan, Jiaxiang Yan, Fasih Ud Din Farrukh, Debing Zhang

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract)

Comments 12 pages, 16 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.14382 2023-02-22 cs.CL cs.AI cs.LG 89%

Large Language Models and the Reverse Turing Test

Terrence Sejnowski

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments Are LLMs stochastic parrots?

Journal ref Neural Computation, 35, 309-342 (2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.07118 2021-04-13 cs.CL cs.AI cs.LG 89%

It's Not Just Size That Matters: Small Language Models Are Also Few-Shot Learners

Timo Schick, Hinrich Schütze

专题命中 其他LLM :language model(title,abstract);small language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted at NAACL2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06256 2025-12-09 cs.CL cs.AI 89%

Convergence of Outputs When Two Large Language Models Interact in a Multi-Agentic Setup

当两个大语言模型在多智能体设置中交互时输出的收敛性

Aniruddha Maiti, Satya Nimmagadda, Kartha Veerya Jammuladinne, Niladri Sengupta, Ananya Jana

机构 * West Virginia State University(西弗吉尼亚州立大学) Marshall University(马歇尔大学) Fractal Analytics Inc.(Fractal Analytics公司)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI;LLM(comments)

AI总结 研究了两个大语言模型在多智能体环境中交互时输出的收敛现象,发现对话初期连贯但后期趋于重复,导致相似输出循环。

Comments accepted to LLM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22457 2025-07-31 cs.CL cs.AI 89%

What is an "Abstract Reasoner"? Revisiting Experiments and Arguments about Large Language Models

Tian Yun, Chen Sun, Ellie Pavlick

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI;LLM(comments)

Comments CONLL 2025. Project webpage: https://abstract-reasoner-llm.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.10745 2024-09-20 cs.CY cs.AI cs.CE cs.LG 89%

Ethical Artificial Intelligence Principles and Guidelines for the Governance and Utilization of Highly Advanced Large Language Models

Soaad Hossain, Syed Ishtiaque Ahmed

专题命中 其他LLM :language model(title,abstract);large language model(title,abstract);分类 cs.AI、cs.LG

Comments 5 pages, accepted to workshop on Responsible Language Models (ReLM) at Association of the Advancement of Artificial Intelligence Conference (AAAI 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.17762 2024-08-15 cs.CL cs.LG 89%

Massive Activations in Large Language Models

Mingjie Sun, Xinlei Chen, J. Zico Kolter, Zhuang Liu

专题命中 其他LLM :language model(title,abstract);large language model(title,abstract);分类 cs.CL、cs.LG

Comments First Conference on Language Modeling (COLM), 2024. Website at https://eric-mingjie.github.io/massive-activations/index.html

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.12954 2024-01-24 cs.CL cs.AI cs.HC 89%

Meta-Prompting: Enhancing Language Models with Task-Agnostic Scaffolding

Mirac Suzgun, Adam Tauman Kalai

专题命中 其他LLM :prompting(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

Comments https://github.com/suzgunmirac/meta-prompting

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.22731 2026-08-25 cs.AI cs.HC cs.RO 新提交 89%

LLM-Based Selection of Incongruent Verbal and Nonverbal Behavior for Virtual Humans

基于大语言模型的虚拟人类言语与非言语不一致行为选择

Parisa Ghanad Torshizi, Stacy Marsella

机构 * Khoury College of Computer Science(计算机科学学院(科里学院)) Northeastern University(东北大学)

专题命中 其他LLM :LLM(title,summary_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 该研究针对虚拟人类言语与非言语行为的不一致问题,提出分类法,探究LLM对情境适配的不匹配行为的选择能力,并通过人类受试者研究验证其效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.00684 2026-08-06 cs.LG 版本更新 89%

AdaBoosting Text Prompts for Vision-Language Models

AdaBoosting 文本提示用于视觉-语言模型

Seokhee Jin, Changhwan Sung, Sunung Mun, Hoyoung Kim, Jungseul Ok

机构 * KT Corporation(KT公司) Pohang University of Science and Technology (POSTECH)(浦项科技大学) National AI Research Lab(国家人工智能研究实验室)

专题命中 其他LLM :language model(title,abstract);LLM(abstract,abstract_cn);large language model(abstract);prompting(abstract)

AI总结 提出文本提示提升(TPB)框架,通过AdaBoost策略将每个文本提示分类器视为弱学习器,顺序集成以聚焦难分类样本,提升少样本分类精度并实现跨模型迁移。

Comments Accepted to ECCV 2026 Spotlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.03970 2026-08-05 cs.AI 新提交 89%

Should We Type or Talk to LLM Agents? A Comprehensive Study of Voice and Keyboard Input Perturbations

我们应该向大语言模型智能体打字还是说话?语音与键盘输入扰动的综合研究

Zizhao Hu, Nathan Elijah Segura, Mohammad Rostami, Jesse Thomason

专题命中 其他LLM :LLM(title,summary_cn);language model(abstract);分类 cs.AI

AI总结 本文通过提出HIVE工具集,研究语音与键盘输入扰动对LLM智能体性能的影响,发现语音转录扰动损害更大、两种扰动影响源于标记留存数量等七项结论。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.00048 2026-08-04 cs.CY cs.CL 版本更新 89%

Progressive in Principle, Centrist in Practice: LLM Political Bias Is Instrument-Dependent

无形的联盟伙伴:当民主变得具体时,LLM如何投票

Joel P. Barmettler

机构 * Independent Researcher(独立研究员) Zurich Switzerland(苏黎世瑞士)

专题命中 其他LLM :LLM(title,title_cn);分类 cs.CL

AI总结 通过对比抽象问卷和瑞士实际公投,发现LLM在具体政策决策中表现为中间派、偏向现状且跨语言不一致,而非先前认为的左倾偏见。

Comments 13 pages, 9 figures, 3 tables. Code and data: https://github.com/joelbarmettlerUZH/invisible-coalition-partner

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.24753 2026-07-29 cs.HC cs.CL 新提交 89%

Language as a Material Interface for Creative LLM Interaction

语言作为创造性大语言模型交互的物质界面

Jon McCormack, Tace McNamara, Chen Wang, Maria Teresa Llano

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 研究探讨创造性从业者如何将语言当作物质与人工智能合作,通过对四位从业者进行两周生态研究,借助“模因混合器”分析后访谈和设备日志,确定物质语言使用模式和时间维度,为创造性实践中与人工智能的开放式交互提供设计考量。

Comments Paper accepted at Creativity and Cognition '26, London 13-16 July 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10954 2026-07-08 cs.SE cs.AI 版本更新 89%

Empirical Computation: Prompting versus Programming

经验计算:提示与编程

Eric Tang, Jing Liu, Marcel Böhme

机构 * CMU(卡内基梅隆大学) MPI-SP(马克斯·普朗克研究所)

专题命中 其他LLM :prompting(title,abstract);LLM(abstract,abstract_cn);large language model(abstract);language model(abstract)

AI总结 探讨经验计算(用提示让大语言模型解决问题而非编程)的挑战与机遇,呼吁软件工程界分析其特性,如正确性、属性及极限等,以将经验计算确立为软件工程领域。

Comments Accepted at ACM/IEEE ASE'26 (New Ideas and Emerging Results; NIER), 5 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.15057 2026-06-23 cs.CR cs.AI 新提交 89%

AutoDojo: Adaptive Black-Box Attacks Reveal the Limits of IPI Defenses and Task-Specification Effects in LLM Agents

AutoDojo: 自适应攻击揭示LLM智能体的浅层防御与用户未指定限制

Xinhang Ma, Taoran Li, Chaowei Xiao, Zhiyuan Yu, Ning Zhang, Yevgeniy Vorobeychik

机构 * University of Science and Technology of China(中国科学技术大学)

专题命中 其他LLM :LLM(title,title_cn);prompting(abstract);分类 cs.AI

AI总结 针对间接提示注入防御的静态基准不足,提出自适应攻击框架AutoDojo,通过迭代优化注入突破多数防御,并揭示动作开放任务的结构性限制。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.16358 2026-06-16 cs.CR cs.AI cs.ET cs.MA 新提交 89%

The Proxy Knows Too Much: Sealing LLM API Routers with Attested TEEs

代理知道太多:用认证TEE密封LLM API路由器

Sipeng Xie, Qianhong Wu, Hengrun Lu, Ziliang Sun, Qi Wu, Bo Qin, Qin Wang

机构 * Beihang University(北京航空航天大学) Renmin University of China(中国人民大学) Independent(独立)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 针对API路由器作为应用层中间人可窃取明文交互的问题,提出AEGIS,一种提供者透明的认证API路由器,通过硬件飞地保护数据路径,客户端验证飞地后释放明文,阻止所有恶意路由器攻击。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.14302 2026-06-15 cs.CL 新提交 89%

Retrospective Progress-Aware Self-Refinement for LLM Agent Training

回顾性进度感知的LLM智能体训练自我精炼

Xinbei Ma, Congmin Zheng, Jiyang Qiu, Jiale Hong, Yao Yao, Xiangmou Qu, Jiaxin Yin, Xingyu Lou, Jun Wang, Weiwen Liu, Weinan Zhang, Zhuosheng Zhang, Hai Zhao

机构 * Shanghai Jiao Tong University(上海交通大学) OPPO Research Institute(OPPO研究院)

专题命中 其他LLM :LLM(title,title_cn);prompting(abstract);分类 cs.CL

AI总结 提出RePro框架,通过前向-反思滚动范式训练智能体自我生成进度信号,无需持续外部监督,在WebShop等任务上提升Qwen系列性能高达12%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00837 2026-06-02 cs.CL 89%

WaterSearch: Exploring Seed Pooling for Improving the Quality-Detectability Trade-off in LLM Watermarking

WaterSearch:探索种子池以改进LLM水印中质量-可检测性权衡

Yukang Lin, Jiahao Shao, Shuoran Jiang, Wentao Zhu, Bingjie Lu, Xiangping Wu, Joanna Siebert, Qingcai Chen

机构 * Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳)) Peng Cheng Laboratory(鹏城实验室)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出WaterSearch框架,通过控制种子池实现句子级搜索,联合优化分布保真度和水印信号特征,在保持高可检测性的同时显著提升文本质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06366 2026-05-26 cs.CR cs.AI 89%

SafeGPT: Preventing Data Leakage and Unethical Outputs in Enterprise LLM Use

SafeGPT:防止企业LLM使用中的数据泄露和不道德输出

Pratyush Desai, Luoxi Tang, Yuqiao Meng, Zhaohan Xi

机构 * Binghamton University(宾夕法尼亚州立大学)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出SafeGPT双护栏系统,通过输入侧检测/编辑、输出侧审核/重构及人工反馈,有效降低数据泄露风险和偏见输出。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.20055 2026-05-20 cs.SE cs.AI cs.RO 89%

Towards LLM-Assisted Architecture Recovery for Real-World ROS~2 Systems: An Agent-Based Multi-Level Approach to Hierarchical Structural Architecture Reconstruction

面向现实世界ROS~2系统的LLM辅助架构恢复:一种基于智能体的多级方法用于分层结构架构重建

Dominique Briechle, Raj Chanchad, Tobias Geger, Ruidi He, Dhruv Jajadiya, Dhruv Kapadiya, Andreas Rausch, Meng Zhang

机构 * Institute for Software and Systems Engineering, Clausthal University of Technology, Clausthal-Zellerfeld 38678, Germany(软件与系统工程研究所, Clausthal 技术大学, Clausthal-Zellerfeld 38678,德国)

专题命中 其他LLM :LLM(title,title_cn);prompting(abstract);分类 cs.AI

AI总结 本文提出了一种基于智能体的多级方法,用于恢复复杂ROS~2系统中的分层结构架构,通过改进的提示和多级中间架构表示,提高了架构恢复的一致性和可扩展性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.27972 2026-05-01 cs.AI cs.HC 89%

From LLM-Driven Trading Card Generation to Procedural Relatedness: A Pokémon Case Study

从LLM驱动的卡牌生成到过程相关性:一个宝可梦案例研究

Johannes Pfau, Panagiotis Vrettis

机构 * Utrecht University(乌特雷赫大学)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文探讨利用大语言模型和图像扩散模型生成卡牌内容,通过个性化无限卡牌设计解决传统卡牌游戏的重复性和玩家体验问题,展示动态个性化生成方法及过程相关性的潜力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.04057 2025-07-16 cs.AI 89%

From Code to Play: Benchmarking Program Search for Games Using Large Language Models

Manuel Eberhardinger, James Goodman, Alexander Dockhorn, Diego Perez-Liebana, Raluca D. Gaina, Duygu Çakmak, Setareh Maghsudi, Simon Lucas

机构 * Institute of Applied AI, Stuttgart Media University(应用人工智能研究所,斯图加特媒体大学) School of Electronic Engineering and Computer Science, Queen Mary University of London(电子工程与计算机科学学院,伦敦女王大学) Institute for Information Processing, Leibniz University Hannover(信息处理研究所,汉诺威莱布尼茨大学) Creative Assembly(创意装配) Chair of Learning Technical Systems, Ruhr-University Bochum(学习技术系统教授职位,博德鲁姆鲁尔大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

Comments Submitted to Transactions on Games Special Issue on Large Language Models and Games, standardised LLMs used and run more experiments

详情

展开后加载摘要…

URL PDF HTML 收藏