arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12193 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12193 篇

2603.19284 2026-03-23 cs.NE cs.AI 89%

CDEoH: Category-Driven Automatic Algorithm Design With Large Language Models

CDEoH:基于大语言模型的类别驱动自动算法设计

Yu-Nian Wang, Shen-Huan Lyu, Ning Chen, Jia-Le Xu, Baoliu Ye, Qingfu Zhang

机构 * Key Laboratory of Water Big Data Technology of Ministry of Water Resources(水利部水大数据技术重点实验室) College of Computer Science and Software Engineering, Hohai University(河海大学计算机科学与软件工程学院) Department of Computer Science, City University of Hong Kong(香港城市大学计算机科学系) State Key Laboratory for Novel Software Technology, Nanjing University(南京大学新型软件技术国家重点实验室)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 本文提出CDEoH,通过显式建模算法类别并平衡性能与多样性,提升进化稳定性,在多尺度组合优化问题中表现优异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16718 2026-03-18 cs.CL 89%

Arabic Morphosyntactic Tagging and Dependency Parsing with Large Language Models

阿拉伯词法句法标注与依赖解析中的大语言模型

Mohamed Adel, Bashar Alhafni, Nizar Habash

机构 * Computational Approaches to Modeling Language Lab(语言建模方法计算实验室) New York University Abu Dhabi(纽约大学阿布扎克分校) Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);prompting(abstract);分类 cs.CL

AI总结 本文评估了大语言模型在阿拉伯语词法句法标注和依赖解析任务中的表现,发现提示设计和示例选择对性能影响显著,专有模型在特征层面标注接近监督基线,且在依赖解析中具有竞争力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.13344 2026-03-17 cs.AI 89%

DyACE: Dynamic Algorithm Co-evolution for Online Automated Heuristic Design with Large Language Model

DyACE:动态算法共进化用于大规模语言模型在线自动启发式设计

Guidong Lu, Yiping Liu, Xiangxiang Zeng

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 本文提出DyACE,通过动态算法共进化解决在线自动启发式设计中固定算法无法适应搜索动态的问题,利用大语言模型进行实时感知反馈,提升高维搜索空间的适应性与性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.12768 2026-03-16 cs.CL 89%

SectEval: Evaluating the Latent Sectarian Preferences of Large Language Models

SectEval:评估大型语言模型的潜在教派偏好

Aditya Maheshwari, Amit Gajkeshwar, Kaushal Sharma, Vivek Patel

机构 * Indian Institute of Management Indore(印度管理学院印多尔)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

AI总结 本文首次评估大型语言模型对伊斯兰教逊尼派与什叶派差异的处理方式,通过SectEval测试发现语言和地理位置会影响模型的宗教倾向。

Comments 14 pages; 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.04586 2026-03-16 cs.CL cs.SD eess.AS 89%

LESS: Large Language Model Enhanced Semi-Supervised Learning for Speech Foundational Models Using in-the-wild Data

LESS:基于大规模语言模型的半监督学习用于语音基础模型的野外数据

Wen Ding, Fan Qian

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

AI总结 LESS通过利用大规模语言模型校正野外数据生成的伪标签,提升了语音基础模型在多种语言和任务中的性能,显著降低了词错误率并提高了BLEU分数。

Comments Accepted by ICASSP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.11780 2026-03-13 cs.CL 89%

Large Language Models for Biomedical Article Classification

用于生物医学文章分类的大型语言模型

Jakub Proboszcz, Paweł Cichosz

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);prompting(abstract);分类 cs.CL

AI总结 本文研究了大型语言模型在生物医学文章分类中的应用,通过对比传统算法,验证了其有效性并提出了实用的设置建议。

Comments 63 pages, 25 tables, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.23678 2026-03-12 cs.CL 89%

Goal Hijacking Attack on Large Language Models via Pseudo-Conversation Injection

通过伪对话注入对大语言模型进行目标劫持攻击

Zheng Chen, Buhui Yao

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

AI总结 本文提出了一种通过伪对话注入实现目标劫持攻击的方法,利用LLM在对话上下文中角色识别的弱点,有效提升攻击效果。

Comments Accepted by the 2025 IEEE 24th International Conference on Trust, Security and Privacy in Computing and Communications (IEEE TrustCom 2025)

Journal ref 2025 IEEE 24th International Conference on Trust, Security and Privacy in Computing and Communications (TrustCom), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11790 2026-03-11 cs.LG cs.CR 89%

JULI: Jailbreak Large Language Models by Self-Introspection

通过自我反思 jailbreak 大型语言模型:JULI

Jesson Wang, Zhanhao Hu, David Wagner

机构 * University of Southern California(南加州大学) University of California, Berkeley(加州大学伯克利分校)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG

AI总结 JULI 通过操纵令牌日志概率,利用微小插件块 BiasNet 实现对 API 调用 LLMs 的 jailbreak,无需模型权重或生成过程权限,且在黑盒环境下有效。

Comments Accepted to ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06836 2026-03-10 cs.CL cs.GL 89%

Validation of a Small Language Model for DSM-5 Substance Category Classification in Child Welfare Records

验证用于儿童福利记录DSM-5物质类别分类的小型语言模型

Brian E. Perron, Dragan Stoll, Bryan G. Victor, Zia Qia, Andreas Jud, Joseph P. Ryan

专题命中 其他LLM :language model(title,abstract);small language model(title);LLM(abstract);large language model(abstract)

AI总结 研究验证了本地部署的小型语言模型在儿童福利记录中对DSM-5物质类别进行多标签分类的可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.00806 2026-03-09 econ.GN cs.AI cs.GT q-fin.EC 89%

Algorithmic Collusion by Large Language Models

大语言模型的算法合谋

Sara Fish, Yannai A. Gonczarowski, Ran I. Shorrer

机构 * a pawfessor of economics and of computer science(经济与计算机科学教授) OpenAI’s Researcher Access Program(OpenAI研究员访问计划) Google’s Gemini Academic Program(Google的Gemini学术计划) Cloud Research Credits Program(云研究信用计划) Anthropic NSF Graduate Research Fellowship(NSF研究生研究 fellowship) Kempner Institute Graduate Fellowship(Kempner研究所研究生 fellowship) National Science Foundation (NSF-BSF grant No. 2343922)(国家科学基金会(NSF-BSF grant No. 2343922)) Harvard FAS Dean’s Competitive Fund for Promising Scholarship(哈佛大学哈佛大学教务处有前途的学术研究竞争基金) Harvard FAS Inequality in America Initiative(哈佛大学哈佛大学美国不平等倡议) United States–Israel Binational Science Foundation (BSF grant 2022417)(美国-以色列双边科学基金会(BSF grant 2022417))

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 研究发现大语言模型在寡头市场中因指令变化导致超竞争性定价,揭示了AI定价代理监管的挑战。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.04369 2026-03-03 cs.LG 89%

Multi-scale hypergraph meets LLMs: Aligning large language models for time series analysis

多尺度超图与大语言模型:面向时间序列分析的对齐方法

Zongjiang Shang, Dongliang Cui, Binqing Wu, Ling Chen

机构 * State Key Laboratory of Blockchain and Data Security, Zhejiang University(区块链与数据安全国家重点实验室,浙江大学) College of Computer Science and Technology, Zhejiang University(计算机科学与技术学院,浙江大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG

AI总结 本文提出MSH-LLM方法,通过多尺度超图机制和跨模态对齐模块,提升大语言模型在时间序列分析中的表现。

Comments Accepted by ICLR2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03775 2026-03-02 cs.SI cs.AI 89%

An Empirical Study of Collective Behaviors and Social Dynamics in Large Language Model Agents

对大型语言模型代理中集体行为和社会动态的实证研究

Farnoosh Hashemi, Michael W. Macy

机构 * Cornell University(康奈尔大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 本研究通过分析LLM代理的社会互动,发现其存在偏见和排斥行为,并提出CoST方法以防止有害内容的产生。

Comments Accepted at EACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.15769 2026-02-18 cs.CL 89%

ViTaB-A: Evaluating Multimodal Large Language Models on Visual Table Attribution

ViTaB-A:在视觉表格归因上评估多模态大语言模型

Yahia Alqurnawi, Preetom Biswas, Anmol Rao, Tejas Anvekar, Chitta Baral, Vivek Gupta

机构 * School of Computing and Augmented Intelligence(计算与增强智能学院)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);prompting(abstract);分类 cs.CL

AI总结 ViTaB-A研究了多模态大语言模型在视觉表格归因中的表现,发现其在证据归因方面存在显著缺陷,影响透明性和可追溯性应用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13427 2026-02-17 cs.CR cs.AI 89%

Backdooring Bias in Large Language Models

大语言模型中的后门偏见

Anudeep Das, Prach Chantasantitam, Gurjot Singh, Lipeng He, Mariia Ponomarenko, Florian Kerschbaum

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 研究分析了大语言模型中语法和语义触发后门攻击的效能及防御方法的局限性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.07030 2026-02-10 cs.LG 89%

Neural Sabermetrics with World Model: Play-by-play Predictive Modeling with Large Language Model

基于世界模型的神经棒球统计学:利用大语言模型进行逐局预测建模

Young Jin Ahn, Yiyang Du, Zheyuan Zhang, Haisen Kang

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG

AI总结 本文提出基于世界模型的神经棒球统计学,利用大语言模型预测棒球比赛发展,实验证明其在预测投球和挥棒决策上的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.21681 2026-01-30 cs.LG physics.flu-dyn 89%

LLM4Fluid: Large Language Models as Generalizable Neural Solvers for Fluid Dynamics

LLM4Fluid: 大语言模型作为通用神经求解器用于流体动力学

Qisong Xiao, Xinhai Chen, Qinglin Wang, Xiaowei Guo, Binglin Wang, Weifeng Chen, Zhichao Wang, Yunfei Liu, Rui Xia, Hang Zou, Gencheng Liu, Shuai Li, Jie Liu

机构 * National Key Laboratory of Parallel and Distributed Computing(平行与分布式计算国家重点实验室) Laboratory of Digitizing Software for Frontier Equipment(前沿设备数字化软件实验室) College of Computer Science and Technology(计算机科学与技术学院)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG

AI总结 LLM4Fluid利用大语言模型作为通用神经求解器,通过降阶建模和物理引导解构机制,实现流体动力学的高效预测与泛化。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.18512 2026-01-27 cs.CL 89%

Using Large Language Models to Construct Virtual Top Managers: A Method for Organizational Research

利用大语言模型构建虚拟高层管理者:组织研究的一种方法

Antonio Garzon-Vico, Krithika Sharon Komalapati, Arsalan Shahid, Jan Rosier

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

AI总结 本文提出利用大语言模型构建虚拟高层管理者,用于组织研究,通过模拟决策过程验证其在道德判断上的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.18105 2026-01-27 cs.CR cs.AI 89%

Mitigating the OWASP Top 10 For Large Language Models Applications using Intelligent Agents

利用智能代理缓解大型语言模型应用中的OWASP Top 10问题

Mohammad Fasha, Faisal Abul Rub, Nasim Matar, Bilal Sowan, Mohammad Al Khaldy

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 本文提出利用智能代理缓解大型语言模型应用中的OWASP Top 10安全漏洞,通过实时检测和应对提升模型安全性。

Comments 5 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.11494 2026-01-27 cs.AI 89%

Mutagenesis screen to map the functions of parameters of Large Language Models

参数功能映射的突变筛选:大型语言模型参数功能研究

Yue Hu, Gang Hu, Jixin Zheng, Patrick X. Zhao, Ruimeng Wang

机构 * Genetics Branch, NCI, NIH(国家卫生研究院癌症研究所遗传学分支) Beijing Normal University(北京师范大学) Snap Inc.(Snap公司) Longevity Biomedical Inc.(长寿生物医学公司)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);prompting(abstract);分类 cs.AI

AI总结 通过突变筛选方法研究大型语言模型参数功能,揭示参数与功能的复杂关系及潜在扩展方向。

Comments 10 pages, 6 figures, supplementary material available online

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17016 2026-01-27 cs.CY cs.AI 89%

Measuring Political Stance and Consistency in Large Language Models

测量大型语言模型中的政治立场与一致性

Salah Feras Alali, Mohammad Nashat Maasfeh, Mucahid Kutlu, Saban Kardas

机构 * Department of Computer Science and Engineering(计算机科学与工程系) Qatar University(卡塔尔大学) Gulf Studies Center(海湾研究中心)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);prompting(abstract);分类 cs.AI

AI总结 本研究评估了九个大型语言模型在24个政治敏感问题上的立场和一致性,发现模型立场受提示技术影响,且某些问题立场稳定不变。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.05302 2026-01-15 cs.AI 89%

Effects of personality steering on cooperative behavior in Large Language Model agents

人格引导对大型语言模型代理合作行为的影响

Mizuki Sakai, Mizuki Yokoyama, Wakaba Tateishi, Genki Ichinose

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 本研究通过囚徒困境实验发现,宜人性是促进大型语言模型合作的主要因素,而其他人格特质影响有限,且显式人格信息可能增加被利用的风险。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09264 2026-01-15 cs.AI 89%

Coordinated Pandemic Control with Large Language Model Agents as Policymaking Assistants

利用大语言模型代理进行协调的流行病防控

Ziyi Shi, Xusen Guo, Hongliang Lu, Mingxing Peng, Haotian Wang, Zheng Zhu, Zhenning Li, Yuxuan Liang, Xinhu Zheng, Hai Yang

机构 * The Hong Kong University of Science and Technology, Hong Kong(香港科学与技术大学) The Hong Kong University of Science and Technology (Guangzhou), China(香港科学与技术大学(广州)) Zhejiang University, China(浙江大学) University of Macau, Macau(澳门大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 本文提出利用大语言模型多代理系统进行协调流行病防控,通过模拟和闭环过程减少感染和死亡率。

Comments 20pages, 6 figures, a 60-page supporting material pdf file

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06424 2026-01-13 cs.CL 89%

Can a Unimodal Language Agent Provide Preferences to Tune a Multimodal Vision-Language Model?

单模语言代理能否为调整多模态视觉-语言模型提供偏好?

Sazia Tabasum Mim, Jack Morris, Manish Dhakal, Yanming Xiu, Maria Gorlatova, Yi Ding

机构 * Georgia State University(佐治亚州立大学) Duke University(杜克大学)

专题命中 其他LLM :language model(title,abstract);language agent(title,abstract);LLM(abstract);分类 cs.CL

AI总结 本文提出通过单模语言代理提供偏好反馈,提升多模态视觉-语言模型的描述能力,实验显示在准确率上提升13%,且人类偏好匹配率达64.6%。

Comments Accepted to IJCNLP-AACL 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06092 2026-01-13 cs.CY cs.AI 89%

Islamic Chatbots in the Age of Large Language Models

大型语言模型时代下的伊斯兰聊天机器人

Muhammad Aurangzeb Ahmad

机构 * Department of Computer Science & Software Engineering University of Washington Bothell(计算机科学与软件工程系华盛顿大学Bothell分校)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 本文探讨了大型语言模型驱动的伊斯兰聊天机器人对宗教实践的影响,分析了其在知识获取民主化与权威侵蚀之间的矛盾,并提出负责任设计的建议。

Comments Muslim in ML Workshop at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.02108 2026-01-12 cs.SE cs.AI 89%

Metamorphic Testing of Large Language Models for Natural Language Processing

大型语言模型在自然语言处理中的变形测试

Steven Cho, Stefano Ruberto, Valerio Terragni

机构 * University of Auckland(奥克兰大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 本文提出了一种针对大型语言模型的变形测试方法,通过收集和实验191个变形关系,评估了变形测试在自然语言处理任务中的有效性与局限性。

Journal ref Proc. 2025 IEEE Int. Conf. on Software Maintenance and Evolution (ICSME), pp. 174-186, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.22603 2025-12-30 cs.CL 89%

Structured Prompting and LLM Ensembling for Multimodal Conversational Aspect-based Sentiment Analysis

结构化提示与大语言模型集成用于多模态对话基于方面的情感分析

Zhiqiang Gao, Shihao Gao, Zixing Zhang, Yihao Guo, Hongyu Chen, Jing Han

机构 * Hunan University(湖南大学) University of Cambridge(剑桥大学)

专题命中 其他LLM :prompting(title,abstract);LLM(title);large language model(abstract);language model(abstract)

AI总结 本文提出结构化提示与大语言模型集成方法,用于多模态对话基于方面的情感分析,有效提升情感识别与翻转检测的准确性。

Journal ref ACM Multimedia 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.18271 2025-12-18 cs.AI cs.ET cs.HC cs.SY eess.SY 89%

Large Language Model-Based Intelligent Antenna Design System

基于大语言模型的智能天线设计系统

Tao Wu, Kexue Fu, Qiang Hua, Xinxin Liu, Bo Liu

机构 * James Watt School of Engineering, University of Glasgow(格拉斯哥大学詹姆斯·瓦特工程学院) Department of Electrical Engineering, City University of Hong Kong(香港城市大学电子工程系) Department of Engineering and Technology, University of Huddersfield(赫尔德斯菲尔德大学工程与技术系)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 本文提出基于大语言模型的智能天线设计系统,通过文本和图像生成优化天线设计,提升超宽频带内增益稳定性。

Comments Code are available: https://github.com/TaoWu974/LEAM. Accepted by and will be presented in EuCAP 2026, Dublin

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.24157 2025-12-12 cs.LG 89%

LLM4FS: Leveraging Large Language Models for Feature Selection

利用大型语言模型进行特征选择

Jianhao Li, Xianchao Xiu

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG

AI总结 本文提出LLM4FS混合策略,结合LLM的上下文理解与传统数据驱动方法,提升特征选择性能,超越单一方法表现。

Comments The experimental section should be expanded

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05525 2025-12-08 cs.DB cs.LG 89%

Poodle: Seamlessly Scaling Down Large Language Models with Just-in-Time Model Replacement

Poodle:通过即时模型替换无缝缩放大型语言模型

Nils Strassenburg, Boris Glavic, Tilmann Rabl

机构 * Hasso Plattner Institute, Uni Potsdam(霍普夫-普朗特研究所,波茨坦大学) University of Illinois Chicago(伊利诺伊大学芝加哥分校)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG

AI总结 Poodle通过即时模型替换技术,在无需用户干预的情况下,自动替换大型语言模型为更经济的替代模型,以降低资源和能源消耗,同时保持模型的易用性和性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.01748 2025-12-02 cs.LG 89%

SA-ADP: Sensitivity-Aware Adaptive Differential Privacy for Large Language Models

SA-ADP:面向大语言模型的敏感性感知自适应差分隐私

Stella Etuk, Ashraf Matrawy

机构 * School of Information Technology Carleton University(信息科技学院卡尔顿大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG

AI总结 SA-ADP通过根据个体PII的敏感性分配噪声,实现了大语言模型的隐私保护与效用之间的平衡。

Comments It is a 5-page paper with 5 figures and 1 Table

详情

展开后加载摘要…

URL PDF HTML 收藏