arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12241 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12241 篇

2511.02866 2026-02-25 cs.SE cs.AI cs.AR cs.CR 83%

LM-Fix: Lightweight Bit-Flip Detection and Rapid Recovery Framework for Language Models

LM-Fix:一种轻量级位翻转检测与快速恢复框架用于语言模型

Ahmad Tahmasivand, Noureldin Zahran, Saba Al-Sayouri, Mohammed Fouda, Khaled N. Khasawneh

机构 * The National Institutes of Health(美国国家卫生研究院)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.AI

AI总结 LM-Fix通过轻量级检测和局部修复技术,有效提升大型语言模型在生产环境中的可靠性。

Comments Accepted at IEEE ICCD 2025. Code: https://github.com/ata990/lm-fix. Detects over 94 percent single-bit flips (near 100 percent multi-bit) with about 1 to 7.7 percent overhead; recovery is over 100x faster than a full reload. Keywords: LLMs, bit-flip, fault injection, reliability, security, Rowhammer, SDC, Jailbreaking, Attack, Defense, GPU DRAM faults

Journal ref Proc. IEEE Int. Conf. on Computer Design (ICCD), 2025, pp. 432-440

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.16684 2026-02-19 cs.LG 83%

Retrieval-Augmented Foundation Models for Matched Molecular Pair Transformations to Recapitulate Medicinal Chemistry Intuition

基于检索增强的基础模型进行匹配分子对变换以重现药物化学直觉

Bo Pan, Peter Zhiping Zhang, Hao-Wei Pang, Alex Zhu, Xiang Yu, Liying Zhang, Liang Zhao

机构 * Department of Computer Science, Emory University(计算机科学系,埃默里大学) Merck & Co., Inc.(默克公司)

专题命中 其他LLM :foundation model(title,abstract);prompting(abstract);分类 cs.LG

AI总结 本文提出MMPT-RAG框架,通过检索增强的方法生成多样化的分子变换,提升药物化学中的可控性和新颖性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03276 2026-02-13 cs.CL 83%

Do language models accommodate their users? A study of linguistic convergence

语言模型是否适应其用户?对语言趋同的研究

Terra Blevins, Susanne Schmalwieser, Benjamin Roth

机构 * Khoury College of Computer Sciences, Northeastern University(东北大学克劳尔计算机科学学院) Faculty of Computer Science, University of Vienna(维也纳大学计算机科学系) Faculty of Philological and Cultural Studies, University of Vienna(维也纳大学语言与文化研究系)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL

AI总结 本文研究语言模型是否趋同于用户语言模式,发现模型在对话风格上显著趋同,但不同模型和设置下趋同程度存在差异。

Comments EACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10566 2026-02-10 cs.LG 83%

ASIDE: Architectural Separation of Instructions and Data in Language Models

ASIDE: 语言模型中指令与数据的架构分离

Egor Zverev, Evgenii Kortukov, Alexander Panfilov, Alexandra Volkova, Soroush Tabesh, Sebastian Lapuschkin, Wojciech Samek, Christoph H. Lampert

机构 * Institute of Science and Technology Austria (ISTA)(奥地利科学与技术研究所) Fraunhofer Heinrich Hertz Institute(弗劳恩霍夫海因里希·赫兹研究所) ELLIS Institute Tübingen(图宾根ELLIS研究所) Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所) Tübingen AI Center(图宾根人工智能中心) Centre of eXplainable Artificial Intelligence(可解释人工智能中心) Technische Universität Berlin(柏林技术大学) Berlin Institute for the Foundations of Learning and Data (BIFOLD)(柏林学习与数据基础研究所)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.LG

AI总结 ASIDE通过在令牌嵌入层面实现指令与数据的分离,提升了语言模型的安全性和性能

Comments ICLR 2026 paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18727 2026-02-10 cs.LG 83%

LogSyn: A Few-Shot LLM Framework for Structured Insight Extraction from Unstructured General Aviation Maintenance Logs

LogSyn: 一种基于少样本学习的LLM框架,用于从非结构化通用航空维护日志中提取结构化洞察

Devansh Agarwal, Maitreyi Chatterjee, Biplab Chatterjee

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 LogSyn通过少样本学习利用LLM将非结构化航空维护日志转化为结构化数据,实现故障模式识别与事件分类,提升航空维护流程和预测分析能力。

Comments Accepted in Proceedings of the 3rd INCOM 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07053 2026-02-06 cs.CL cs.SD eess.AS 83%

TASTE: Text-Aligned Speech Tokenization and Embedding for Spoken Language Modeling

TASTE: 用于语音语言建模的文本对齐语音标记化与嵌入

Liang-Hsuan Tseng, Yi-Chang Chen, Kuan-Yi Lee, Da-Shan Shiu, Hung-yi Lee

机构 * MediaTek Research(联发科技研究) Graduate Institute of Communication Engineering, National Taiwan University(国立台湾大学通信工程研究所) Artificial Intelligence Center of Research Excellence, National Taiwan University(国立台湾大学卓越人工智能研究中心)

专题命中 其他LLM :language model(title,abstract);LLM(abstract);分类 cs.CL

AI总结 TASTE通过语音重建目标实现文本对齐的语音标记化与嵌入,提升语音语言建模的性能。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.22585 2026-02-02 cs.DC cs.LG 83%

HetCCL: Accelerating LLM Training with Heterogeneous GPUs

HetCCL: 通过异构GPU加速大语言模型训练

Heehoon Kim, Jaehwan Lee, Taejeoung Kim, Jongwon Park, Jinpyo Kim, Pyongwon Suh, Ryan H. Choi, Sangwoo Lee, Jaejin Lee

机构 * snu-cse(首尔国立大学计算机科学与工程系) snu-ds(首尔国立大学数据科学研究院) moreh(Moreh公司) samsung(三星研究)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 HetCCL通过统一异构GPU的集体通信库,实现跨供应商的高效训练,提升异构环境下的大语言模型训练性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.14803 2026-01-08 cs.CY cs.AI 83%

OnlineMate: An LLM-Based Multi-Agent Companion System for Cognitive Support in Online Learning

OnlineMate: 基于大语言模型的多智能体伴侣系统用于在线学习中的认知支持

Xian Gao, Zongyun Zhang, Ting Liu, Yuzhuo Fu

机构 * Shanghai Jiao Tong University(上海交通大学)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 OnlineMate通过基于LLM的多智能体系统,结合理论思维,为在线学习提供个性化认知支持,提升学习深度与情感参与。

Comments work in progress

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14435 2026-01-06 cs.CY cs.AI 83%

Choosing a Model, Shaping a Future: Comparing LLM Perspectives on Sustainability and its Relationship with AI

选择模型,塑造未来:比较LLM对可持续性和其与AI关系的视角

Annika Bush, Meltem Aksoy, Markus Pauly, Greta Ontrup

机构 * Research Center Trustworthy Data Science and Security, University Alliance Ruhr(可信数据科学与安全研究中心,鲁尔大学联盟) Department of Computer Science, Technical University Dortmund(计算机科学系,图林根技术大学) Chair of Mathematical Statistics and Applications in Industry, Technical University Dortmund(工业数学统计与应用教授职位,图林根技术大学) Department of Computer Science, University of Duisburg-Essen(计算机科学系,杜伊斯堡-埃森大学)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本研究比较了五个先进LLM对可持续性与AI关系的视角,发现模型间存在显著差异,强调模型选择对可持续性战略的影响。

Comments Accepted for EMNLP Conference

Journal ref Conference on Empirical Methods in Natural Language Processing (EMNLP 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10904 2025-12-29 cs.CR cs.AI 83%

CEKER: A Generalizable LLM Framework for Literature Analysis with a Case Study in Unikernel Security

CEKER:一种通用的LLM框架用于文献分析及在unikernel安全领域的案例研究

Alex Wollman, John Hastings

机构 * The Beacom College of Computer and Cyber Sciences(计算机与网络安全科学学院) Dakota State University(达科他州立大学)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 CEKER通过LLM框架实现文献分析自动化,应用于unikernel安全领域,揭示了攻击面减少及安全缺口,强调动态安全调整的重要性。

Comments 7 pages, 2 figures

Journal ref International Symposium on Intelligent Computing and Networking 2025 (ISICN 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.19950 2025-12-24 cs.CL cs.HC 83%

Bias Beneath the Tone: Empirical Characterisation of Tone Bias in LLM-Driven UX Systems

音调下的偏见:LLM驱动的UX系统中音调偏见的实证刻画

Heet Bodara, Md Masum Mushfiq, Isma Farah Siddiqui

机构 * Faculty of Information Technology, Monash University(信息技术学院,莫纳什大学)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文研究了LLM驱动UX系统中的音调偏见问题,通过合成数据集和弱监督方法,揭示了模型在对话中存在系统性的音调偏差,为设计公平可信的对话AI提供了依据。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12706 2025-12-16 cs.AI cs.SE 83%

Synergizing Code Coverage and Gameplay Intent: Coverage-Aware Game Playtesting with LLM-Guided Reinforcement Learning

协同代码覆盖率与游戏意图:基于LLM引导强化学习的覆盖感知游戏测试

Enhong Mu, Minami Yoda, Yan Zhang, Mingyue Zhang, Yutaka Matsuno, Jialong Li

机构 * College of Computer and Information Science, Southwest University(西南大学计算机与信息科学学院) College of Science and Technology, Nihon University(日本立命馆大学科学技术学院) Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院) Waseda Institute for Advanced Study, Waseda University(早稻田大学高级研究机构)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 SMART框架通过LLM引导强化学习,结合结构验证与功能验证,提升游戏更新测试的覆盖率和任务完成率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12630 2025-12-16 cs.HC cs.AI 83%

ORIBA: Exploring LLM-Driven Role-Play Chatbot as a Creativity Support Tool for Original Character Artists

ORIBA:探索基于大语言模型的对话机器人作为原创角色艺术家创造力支持工具

Yuqian Sun, Xingyu Li, Shunyu Yao, Noura Howell, Tristan Braud, Chang Hee Lee, Ali Asadipour

机构 * Computer Science Research Centre, Royal College of Art(皇家艺术学院计算机科学研究中心) Digital Media, School of Literature, Media, and Communication, Georgia Institute of Technology(佐治亚理工学院数字媒体系) Princeton University(普林斯顿大学) Digital Media, Georgia Institute of Technology(佐治亚理工学院数字媒体系) Division of Integrative Systems and Design, The Hong Kong University of Science and Technology(香港科学大学整合系统与设计 division) Industrial Design Department, College of Engineering, KAIST(韩国科学技术院工程学院工业设计系)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 ORIBA通过大语言模型支持原创角色艺术家的创意过程,平衡AI辅助与创意自主权。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08951 2025-12-11 cs.NE cs.AI cs.GR 83%

AI Co-Artist: A LLM-Powered Framework for Interactive GLSL Shader Animation Evolution

AI共艺术家:一种基于大语言模型的交互式GLSL着色器动画进化的框架

Kamer Ali Yuksel, Hassan Sawaf

机构 * aiXplain Inc.(aiXplain公司)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 AI Co-Artist利用大语言模型降低GLSL着色器创作门槛,通过直观交互实现视觉艺术的进化与生成。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07312 2025-12-09 cs.AR cs.AI cs.DC 83%

DCO: Dynamic Cache Orchestration for LLM Accelerators through Predictive Management

DCO: 通过预测管理实现LLM加速器的动态缓存编排

Zhongchun Zhou, Chengtao Lai, Yuhang Gu, Wei Zhang

机构 * Department of Electronic and Computer Engineering, The Hong Kong University of Science and Technology(电子与计算机工程系,香港科学与技术大学) School of Electronic Science and Engineering, Southeast University(电子科学与工程学院,东南大学)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 DCO通过预测管理实现LLM加速器的动态缓存编排,利用数据流信息优化缓存替换和旁路决策,提升性能至1.8倍,面积仅0.064mm²。

Comments \c{opyright} 2025 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14070 2025-11-27 cs.CR cs.AI 83%

Special-Character Adversarial Attacks on Open-Source Language Model

特殊字符对抗攻击对开源语言模型的攻击

Ephraiem Sarabamoun

机构 * University of Virginia(弗吉尼亚大学)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.AI

AI总结 本文研究了针对开源语言模型的特殊字符对抗攻击,评估了多种攻击方式并揭示了模型的安全漏洞及失败模式。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20627 2025-11-26 cs.AI 83%

Fighting AI with AI: Leveraging Foundation Models for Assuring AI-Enabled Safety-Critical Systems

用AI对抗AI:利用基础模型确保AI赋能的安全关键系统

Anastasia Mavridou, Divya Gopinath, Corina S. Păsăreanu

机构 * KBR Inc.(KBR公司) NASA Ames(美国国家航空航天局阿姆斯研究中心)

专题命中 其他LLM :foundation model(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出利用AI技术解决安全关键系统中AI保证问题,通过REACT和SemaLens两个组件实现需求工程与感知系统验证。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18194 2025-11-25 cs.CL 83%

Agent-as-a-Graph: Knowledge Graph-Based Tool and Agent Retrieval for LLM Multi-Agent Systems

Agent-as-a-Graph: 基于知识图谱的LLM多智能体系统中的工具与智能体检索

Faheem Nizar, Elias Lumer, Anmol Gulati, Pradeep Honaganahalli Basavaraju, Vamse Kumar Subbiah

机构 * Commercial Technology and Innovation Office, PricewaterhouseCoopers(普华永道商业技术与创新办公室)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出基于知识图谱的智能体检索方法,通过构建智能体与工具的关系图谱,提升多智能体系统中工具和智能体的检索效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.15478 2025-11-25 cs.CL 83%

Red Teaming Multimodal Language Models: Evaluating Harm Across Prompt Modalities and Models

针对多模态语言模型的红队测试:评估不同提示模态和模型的有害性

Madison Van Doren, Casey Ford

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL

AI总结 本研究通过红队测试评估多模态语言模型在不同提示模态下的安全性,发现Pixtral 12B的有害响应率最高,而Claude Sonnet 3.5最安全,凸显了建立多模态安全基准的必要性。

Journal ref AAAI 2026 AIGOV Workshop and EurIPS 2025 Workshop on Unifying Perspectives on Learning Biases

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08535 2025-11-12 cs.CV cs.AI 83%

Large Sign Language Models: Toward 3D American Sign Language Translation

Sen Zhang, Xiaoxiao He, Di Liu, Zhaoyang Xia, Mingyu Zhao, Chaowei Tan, Vivian Li, Bo Liu, Dimitris N. Metaxas, Mubbasir Kapadia

机构 * Rutgers University(罗格斯大学) Meta Reality Labs(Meta现实实验室) Qualcomm(高通公司) Walmart Global Tech(沃尔玛全球技术) Roblox PRISMS

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00854 2025-11-04 cs.CL 83%

TriCon-Fair: Triplet Contrastive Learning for Mitigating Social Bias in Pre-trained Language Models

Chong Lyu, Lin Li, Shiqing Wu, Jingling Yuan

机构 * School of Computer Science(计算机科学学院) Artificial Intelligence, Wuhan University of Technology, Wuhan, China(武汉理工大学人工智能学院) Faculty of Data Science, City University of Macau, Macau, China(澳门城市大学数据科学学院)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27016 2025-11-03 cs.CL 83%

Semantically-Aware LLM Agent to Enhance Privacy in Conversational AI Services

Jayden Serenari, Stephen Lee

机构 * Department of Computer Science University of Pittsburgh Pittsburgh, Pennsylvania, USA(计算机科学系 纽约大学 伯利恒,宾夕法尼亚州,美国)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted to IEEE Big Data 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.23271 2025-10-28 cs.CL 83%

Mubeen AI: A Specialized Arabic Language Model for Heritage Preservation and User Intent Understanding

Mohammed Aljafari, Ismail Alturki, Ahmed Mori, Yehya Kadumi

专题命中 其他LLM :language model(title,abstract);prompting(abstract);分类 cs.CL

Comments 21 pages, 2 figures, 3 tables. Includes appendices on ethical guidelines and training framework. Submitted September 04, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.08222 2025-10-28 cs.CV cs.AI cs.CY cs.HC 83%

Refusal as Silence: Gendered Disparities in Vision-Language Model Responses

Sha Luo, Sang Jung Kim, Zening Duan, Kaiping Chen

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.04510 2025-10-24 cs.CL 83%

Heterogeneous Swarms: Jointly Optimizing Model Roles and Weights for Multi-LLM Systems

Shangbin Feng, Zifeng Wang, Palash Goyal, Yike Wang, Weijia Shi, Huang Xia, Hamid Palangi, Luke Zettlemoyer, Yulia Tsvetkov, Chen-Yu Lee, Tomas Pfister

机构 * University of Washington(华盛顿大学) Google Cloud AI Research(谷歌云人工智能研究)

专题命中 其他LLM :LLM(title,abstract);language model(abstract);分类 cs.CL

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.08164 2025-10-21 cs.LG 83%

BLUR: A Bi-Level Optimization Approach for LLM Unlearning

Hadi Reisizadeh, Jinghan Jia, Zhiqi Bu, Bhanukiran Vinzamuri, Anil Ramakrishna, Kai-Wei Chang, Volkan Cevher, Sijia Liu, Mingyi Hong

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.06911 2025-10-09 cs.AI 83%

LLM-Assisted Modeling of Semantic Web-Enabled Multi-Agents Systems with AJAN

Hacane Hechehouche, Andre Antakli, Matthias Klusch

机构 * Daimler Buses GmbH(戴姆勒巴士公司) German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25286 2025-10-07 cs.CY cs.AI 83%

Artificial Authority: From Machine Minds to Political Alignments. An Experimental Analysis of Democratic and Autocratic Biases in Large-Language Models

Natalia Ożegalska-Łukasik, Szymon Łukasik

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23158 2025-10-07 cs.CL 83%

User Feedback in Human-LLM Dialogues: A Lens to Understand Users But Noisy as a Learning Signal

Yuhan Liu, Michael J. Q. Zhang, Eunsol Choi

机构 * New York University(纽约大学)

专题命中 其他LLM :LLM(title,abstract);language model(abstract);分类 cs.CL

Comments EMNLP camera-ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.00714 2025-10-07 cs.SE cs.AI 83%

RFCAudit: An LLM Agent for Functional Bug Detection in Network Protocols

Mingwei Zheng, Chengpeng Wang, Xuwei Liu, Jinyao Guo, Shiwei Feng, Xiangyu Zhang

机构 * Department of Computer Science, Purdue University(计算机科学系,普渡大学)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏