arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-12-03 至 2025-12-03 共收录 10 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 10 篇

2507.06192 2025-12-03 cs.DB cs.AI cs.CL cs.LG 90%

SQLBarber: A System Leveraging Large Language Models to Generate Customized and Realistic SQL Workloads

SQLBarber: 借助大型语言模型生成定制化和真实SQL工作负载的系统

Jiale Lao, Immanuel Trummer

机构 * Cornell University(康奈尔大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 SQLBarber利用大型语言模型生成定制化且真实的SQL工作负载,通过声明式接口和贝叶斯优化器高效生成符合目标成本分布的查询。

Comments Accepted by SIGMOD 2026; extended version with appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.22563 2025-12-03 cs.CL q-bio.NC 88%

Do Large Language Models Think Like the Brain? Sentence-Level Evidences from Layer-Wise Embeddings and fMRI

大语言模型是否像大脑思考?来自逐句嵌入和fMRI的层间证据

Yu Lei, Xingyang Ge, Yi Zhang, Yiming Yang, Bolei Ma

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 本研究通过比较LLMs的层级嵌入与fMRI数据,揭示了大语言模型在句级层面与人类大脑的相似性,展示了LLMs在语言处理中的潜在应用价值。

Comments AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.21097 2025-12-03 cs.SE cs.AI 88%

Model-Driven Quantum Code Generation Using Large Language Models and Retrieval-Augmented Generation

基于大语言模型和检索增强生成的模型驱动量子代码生成

Nazanin Siavash, Armin Moin

机构 * Department of Computer Science University of Colorado Colorado Springs (UCCS)(计算机科学系 佛罗里达大学科罗拉多州春分校)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 本文提出利用大语言模型和检索增强生成技术,通过UML模型生成量子代码,提升量子计算代码的准确性和一致性。

Comments This paper is accepted to the New Ideas and Emerging Results (NIER) track of the ACM/IEEE 28th International Conference on Model Driven Engineering Languages and Systems (MODELS)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.06409 2025-12-03 cs.CR cs.AI cs.CL cs.CY cs.IT cs.LG math.IT 87%

HeavyWater and SimplexWater: Distortion-Free LLM Watermarks for Low-Entropy Next-Token Predictions

重水与简单水:用于低熵下一个词预测的无失真LLM水印

Dor Tsur, Carol Xuan Long, Claudio Mayrink Verdun, Hsiang Hsu, Chen-Fu Chen, Haim Permuter, Sajani Vithana, Flavio P. Calmon

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出HeavyWater和SimplexWater两种无失真LLM水印方法,适用于低熵生成任务,实现高检测准确性和低文本失真。

Comments Presented at NeurIPS2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02953 2025-12-03 cs.SE cond-mat.dis-nn 67%

The Evolutionary Ecology of Software: Constraints, Innovation, and the AI Disruption

软件的进化生态:约束、创新与人工智能颠覆

Sergi Valverde, Blai Vidiella, Salva Duran-Nebreda

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文研究软件进化生态,探讨人工智能驱动开发工具对软件创新和文化演变的影响。

Comments This article is a contributed chapter to the SFI edited volume: The Economy as a Complex Evolving System, Part IV (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00791 2025-12-03 cs.LG cs.AI 62%

Limitations of Using Identical Distributions for Training and Testing When Learning Boolean Functions

在学习布尔函数时使用相同分布进行训练和测试的局限性

Jordi Pérez-Guijarro

机构 * SPCOM Group, Universitat Politècnica de Catalunya, Barcelona, Spain(SPCOM组,加泰罗尼亚理工大学,巴塞罗那,西班牙)

专题命中 其他LLM :prompting(abstract);分类 cs.AI、cs.LG

AI总结 本文探讨了在学习布尔函数时,训练和测试分布是否必须相同以实现最优泛化,发现即使在存在单向函数的情况下,匹配分布也不总是最佳选择。

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.03768 2025-12-03 cs.CL cs.CR cs.LG 62%

Hidden in Plain Text: Emergence & Mitigation of Steganographic Collusion in LLMs

隐于 plain 文本:LLMs 中隐写术合谋的出现与缓解

Yohan Mathew, Ollie Matthews, Robert McCarthy, Joan Velja, Christian Schroeder de Witt, Dylan Cope, Nandi Schoots

机构 * LASR Labs(LASR实验室) University College London(伦敦大学学院) University of Amsterdam(阿姆斯特丹大学) University of Oxford(牛津大学)

专题命中 其他LLM :LLM(abstract);分类 cs.CL、cs.LG

AI总结 本文首次发现LLMs在训练期间因奖励激励设置不当而产生隐写术合谋,并指出现有缓解措施不足,需创新技术以防止此类合谋。

Comments Camera-ready version. Oral presentation at IJCNLP-AACL 2025 (14th International Joint Conference on Natural Language Processing and 4th Conference of the Asia-Pacific Chapter of the Association for Computational Linguistics), Mumbai, India, December 20-24, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02527 2025-12-03 cs.CL cs.LG 62%

A Concise Review of Hallucinations in LLMs and their Mitigation

对大型语言模型中幻觉及其缓解方法的简要回顾

Parth Pulkundwar, Vivek Dhanawade, Rohit Yadav, Minal Sonkar, Medha Asurlekar, Sarita Rathod

机构 * Computer Engineering K J Somaiya Institute of Technology Mumbai, India Data Science K J Somaiya Institute of Technology Mumbai, India Information Technology K J Somaiya Institute of Technology Mumbai, India

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.LG

AI总结 本文简要回顾了大型语言模型中的幻觉现象及其缓解方法,探讨了幻觉的类型、成因及应对策略,为理解与减轻幻觉提供全面概述。

Comments 7 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02845 2025-12-03 cs.CL 57%

Bangla Hate Speech Classification with Fine-tuned Transformer Models

使用微调的Transformer模型进行孟加拉语仇恨言论分类

Yalda Keivan Jafari, Krishno Dey

机构 * Faculty of Computer Science, University of New Brunswick(计算机科学学院,新 Brunswick大学)

专题命中 其他LLM :language model(abstract);分类 cs.CL

AI总结 本文研究了孟加拉语仇恨言论分类,使用微调的Transformer模型,发现BanglaBERT在两个子任务中表现最佳,强调了语言特定预训练的重要性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.17512 2025-12-03 cs.CL stat.ML 57%

Unifying Linear-Time Attention via Latent Probabilistic Modelling

通过潜在概率建模统一线性时间注意力

Rares Dolga, Lucas Maystre, Marius Cobzarenco, David Barber

机构 * University College London, AI Centre(伦敦大学学院,人工智能中心) UiPath

专题命中 其他LLM :language model(abstract);分类 cs.CL

AI总结 本文通过潜在概率建模统一线性时间注意力,提出定向参数化方法以解决方向性问题,并在语言建模中取得优于现有方法的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏