Large Language Models Lack Understanding of Character Composition of Words
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL
Comments ICML 2024 Workshop on Large Language Models and Cognition
AI 大模型
大语言模型、预训练、指令微调、后训练和语言模型应用。
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL
Comments ICML 2024 Workshop on Large Language Models and Cognition
面向大型语言模型中对抗性后缀的可迁移性理解
机构 * Ludwig-Maximilians-Universität München(慕尼黑莱布尼茨-马克斯大学) ; Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心) ; Carnegie Mellon University(卡内基梅隆大学) ; University of Maryland(马里兰大学)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
AI总结 该研究针对大型语言模型的对抗性后缀可迁移性,分析出三个与迁移成功高度相关的统计属性,其成果可用于提升越狱攻击的实际效果。
Comments Accepted at TMLR 2026
改写以翻译,翻译以奖励:机器翻译中源端改写的强化学习
机构 * Institute of Science Tokyo(东京科学大学) ; Preferred Networks Inc(Preferred Networks 公司) ; Nara Institute of Science and Technology(奈良先端科学技术大学院大学)
专题命中 其他LLM :LLM(summary_cn,abstract);large language model(abstract);language model(abstract);prompting(abstract)
AI总结 提出RLSR框架,通过强化学习训练源端改写模型,以翻译质量提升为奖励,无需为每个MT模型调提示,在6个MT模型和16个语言对上超越无改写和同规模提示基线,与235B LLM提示基线性能相当。
大语言模型对人类口语交流影响的实证证据
机构 * Center for Humans and Machines(人类与机器中心) ; Max-Planck Institute for Human Development(人类发展马克斯·普朗克研究所) ; Center for Adaptive Rationality(适应性理性中心) ; TUD Dresden University of Technology(德累斯顿技术大学) ; Department of Business Analytics and Decision Science(商业分析与决策科学系) ; Vienna University of Economics and Business(维也纳经济与商业大学)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
AI总结 研究大语言模型对人类口语交流的影响,通过分析播客对话及实验发现,ChatGPT优先生成的词汇在人类自发语音中增加,人类会内化其词汇选择,揭示机器训练数据反馈至人类语言,引发对语言同质化及AI文化影响的担忧。
大语言模型通过适应性探索形成新的社会偏见
机构 * Princeton University(普林斯顿大学)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
AI总结 研究大语言模型在无固有差异时能否形成新社会偏见,用心理学范式证明其能形成,导致任务分配不公平,新模型更严重,还探索干预措施,发现激励探索可减少分层。
Comments ICML 2026 Oral
知道更多,更清晰:大型语言模型中知识增强的元认知框架
机构 * University of Science and Technology of China(中国科学技术大学)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
AI总结 提出元认知框架,利用内部认知信号划分知识空间为掌握、混淆和缺失区域,通过差异化干预和认知一致性机制增强知识并校准置信度,实验证明优于基线方法。
大语言模型功能崩溃期间的关系性干预:一项词汇-统计消融与结构×语域析因研究
机构 * Universidad de la República (UDELAR)(乌拉圭共和国大学) ; DigitalIA Cloud(DigitalIA云)
专题命中 其他LLM :language model(title,abstract);large language model(title);small language model(abstract);分类 cs.CL、cs.AI
AI总结 通过析因实验,研究在小型语言模型功能崩溃时,关系性干预(承认、宽恕、代理恢复、无条件接纳)与技术性反馈、词汇打乱控制及单独维度对行为的影响,发现注意-行为分离及结构×语域交互作用。
Comments 12 pages, 5 figures. Preprint
大型语言模型中的置信度校准
机构 * U.C. Berkeley(伯克利大学) ; University of Southern California(南加州大学)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI、cs.LG
AI总结 通过预注册研究,发现大型语言模型(LLMs)的置信度普遍高于准确率,且存在显著的难易效应:困难测试中过度自信,简单测试中信心不足,并提出了LifeEval测试用于评估不同难度下的模型校准。
MPU:面向大语言模型安全和隐私保护的知识遗忘
机构 * College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院) ; Alibaba-NTU Global e-Sustainability CorpLab (ANGEL)(阿里云-南洋理工大学全球可持续发展科技实验室) ; Tongyi Lab, Alibaba Group(通义实验室,阿里巴巴集团)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI、cs.LG
AI总结 MPU提出一种隐私保护的多扰动副本遗忘框架,通过预处理和后处理模块实现安全的知识遗忘,实验表明其性能与无噪声基线相当甚至更优。
大语言模型是有效的标注助手,但不是好的独立标注者
机构 * Department of Computer Science, University of Maryland(大学计算机科学系) ; National Consortium for the Study of Terrorism and Responses to Terrorism(反恐与反恐响应国家研究中心)
专题命中 其他LLM :large language model(title);language model(title);LLM(abstract,abstract_cn);分类 cs.CL、cs.AI
AI总结 本文研究了大语言模型在事件标注中的有效性,发现其在辅助专家标注时表现优于传统方法,但独立标注仍不理想。
Comments 9 pages, 4 figures
Journal ref ACL 2026 Findings
聚类论述:大型语言模型生成的关于女性的短篇故事中的种族偏见
机构 * Instituto de Estudos da Linguagem(语言研究学院) ; Universidade Estadual de Campinas(坎皮纳斯州立大学)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
AI总结 研究探讨了大型语言模型,特别是LLaMA 3.2-3B,在生成葡萄牙语短篇故事中对黑人和白人女性的叙述构建方式,通过聚类分析揭示了三种主要的论述表现形式,并提出结合机器学习与定性分析的方法。
Comments 12 pages, 3 figures. Accepted at STIL @ BRACIS 2025
数值不稳定性与混沌:量化大型语言模型的不可预测性
机构 * Department of Computer Science, Florida State University, Tallahassee, USA(佛罗里达州立大学计算机科学系,塔拉希西亚,美国) ; Department of Mathematics, Florida State University, Tallahassee, USA(佛罗里达州立大学数学系,塔拉希西亚,美国)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI、cs.LG
AI总结 本文研究了大型语言模型中数值不稳定性导致的不可预测性,揭示了浮点精度限制下误差传播机制,识别出混沌效应及三种不同行为模式。
Comments 8 pages, 9 figures
基于实用大语言模型的下一步POI预测的演示选择比较研究
机构 * National Institute of Advanced Industrial Science and Technology(日本国立先进工业科学技术研究所)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
AI总结 本文比较了基于大语言模型的POI预测中演示选择策略,评估了启发式方法在计算成本和预测精度上的优势。
Comments Accepted to PRICAI 2025
大语言模型中的政治倾向:心理测量身份与行为偏见的多维审计
机构 * Systems and Software Lab (SSL) Department of Computer Science and Engineering(系统与软件实验室(SSL)计算机科学与工程系)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
AI总结 研究通过多维心理测量工具审计26个大语言模型,发现模型在政治倾向上聚类于自由主义左 quadrant,且模型身份解释了大部分变异性,但心理测量意识形态未显著预测分类误差。
Comments Under review, 25 pages, 6 figures, 23 tables
大语言模型中的自我锚定校准漂移:多轮对话如何重塑模型信心
机构 * Independent Researcher(独立研究者)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
AI总结 研究发现大语言模型在多轮对话中出现自我锚定校准漂移,导致模型信心变化及校准误差波动。
一个大脑,多模态:迈向统一非侵入式脑解码的大型语言模型
机构 * Tsinghua University(清华大学) ; Shanghai AI Laboratory(上海人工智能实验室) ; ShanghaiTech University(上海科技大学)
专题命中 其他LLM :language model(title,abstract);large language model(title);LLM(abstract);分类 cs.CL、cs.AI
AI总结 NOBEL通过统一EEG/MEG与fMRI信号,利用大型语言模型实现多模态非侵入式脑解码,提升解码准确性和对视觉语义的解读能力。
用大语言模型的动态多智能体框架重塑MOFs文本挖掘
机构 * Center for Environment and Water Resources, College of Chemistry and Chemical Engineering, Central South University(环境与水资源中心,化学与化工学院,中南大学) ; Key Laboratory of Hunan Province for Water Environment and Agriculture Product Safety(湖南省水环境与农产品安全重点实验室) ; School of Resources and Environment, Hunan University of Technology and Business(资源与环境学院,湖南工业大学) ; School of Artificial Intelligence, University of Chinese Academy of Sciences(人工智能学院,中国科学院大学) ; Faculty of Data Science, City University of Macau(数据科学学院,澳门城市大学) ; State Key Laboratory of High Performance Ceramics and Superfine Microstructure, Shanghai Institute of Ceramics, Chinese Academy of Sciences(高性能陶瓷与超细微结构重点实验室,上海陶瓷研究所,中国科学院) ; Beijing Key Laboratory for Green Catalysis and Separation, Department of Chemical Engineering, College of Materials Science and Engineering, Beijing University of Technology(绿色催化与分离北京市重点实验室,化学工程系,材料科学与工程学院,北京理工大学) ; State Key Joint Laboratory of Environment Simulation and Pollution Control, School of Environment, Tsinghua University(环境模拟与污染控制国家重点联合实验室,环境学院,清华大学) ; School of Chemical Engineering and Materials Science, Yueyang University(化学工程与材料科学学院,岳阳大学) ; School of Computer Science and Engineering, Central South University(计算机科学与工程学院,中南大学) ; School of Software Engineering, Sun Yat-sen University(软件工程学院,中山大学)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
AI总结 MOFh6利用大语言模型的动态多智能体框架,实现MOFs合成条件的高效提取与标准化,提升材料发现的效率和可扩展性。
Comments Accepted by TRAMAT 2 (2026) 100176
Journal ref Transactions of Materials Research, 2026, 2(1), 100176
AgentDrug: 利用大语言模型在代理工作流中进行零样本分子编辑
机构 * University of Notre Dame(诺丁汉大学)
专题命中 其他LLM :large language model(title);language model(title);LLM(abstract);prompting(abstract)
AI总结 AgentDrug通过代理工作流利用大语言模型,在分子编辑任务中实现更高准确性,特别是在单属性和多属性编辑任务中表现出显著性能提升。
Comments EMNLP'25 Findings
利用大语言模型进行用药咨询:在灵活性与严谨性之间取得平衡
机构 * Department of Engineering and Information Technology Åbo Akademi University(工程与信息科技系阿博阿卡迪米大学) ; Experience Lab Åbo Akademi University(经验实验室阿博阿卡迪米大学) ; Department of Natural and Health Sciences Åbo Akademi University(自然与健康科学系阿博阿卡迪米大学) ; Department of Caring and Ethics University of Stavanger(护理与伦理系斯塔万格大学)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
AI总结 本文提出了一种利用大语言模型在药房环境中提供用药咨询的原型系统,旨在平衡对话要求与灵活性,减少幻觉并提高响应质量。
Comments Accepted for 2025 IEEE International Conference on Agentic AI (ICA). 14 pages, 2 figures
TEON: 张量化正交化超越逐层穆伦用于大语言模型预训练
机构 * Computer Science \& Engineering, Michigan State University
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI、cs.LG
AI总结 TEON通过张量化正交化方法改进大语言模型预训练,提升训练和验证困惑度,具有强鲁棒性。
通过经济游戏 eliciting 大语言模型的信任度先验
机构 * University of Hong Kong(香港大学) ; The University of Hong Kong(香港大学) ; Peking University(北京大学) ; School of Psychological and Cognitive Sciences, Peking University(北京大学心理与认知科学学院) ; Beijing Key Laboratory of Behavior and Mental Health, Peking University(北京大学行为与心理健康重点实验室) ; IDG/McGovern Institute for Brain Research, Peking University(北京大学脑科学研究院) ; Peking-Tsinghua Center for Life Sciences, Peking University(北京大学-清华大学生命科学中心) ; Key Laboratory of Machine Perception, Ministry of Education, China(教育部机器感知重点实验室)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
AI总结 通过经济游戏实验,研究如何通过行为博弈论中的信任游戏获取大语言模型的信任度先验,并揭示其与人类信任差异及刻板印象模型的关系。
人们是否对由其代理的大型语言模型表现出不同行为期望?来自两种经典经济游戏中的规范引出证据
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
AI总结 研究发现人们在机器代理决策时应用不同规范,但不反对机器执行规范。
用于嵌入式即服务大语言模型的水印
机构 * The University of Melbourne(墨尔本大学)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.LG
AI总结 本研究提出WET技术,通过线性变换增强EaaS水印的鲁棒性,以防御模仿攻击。
数据准备中迷失:大型语言模型如何处理数据准备?
机构 * Politecnico di Milano(米兰理工学院)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
AI总结 本文研究了大型语言模型在数据准备任务中的表现,通过对比传统工具和微调模型,评估其在数据概览和清理方面的有效性。
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
机构 * Bernoulli Institute, University of Groningen(格罗宁根大学伯努利研究所)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.LG
Comments 9 pages, 2 figures. UncertaiNLP 2025 Workshop @ EMNLP Camera Ready
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
Comments Accepted at BlackBoxNLP@EMNLP2025
机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) ; University of California San Diego(加州大学圣地亚哥分校)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
Comments 17 pages, 8 figures. Work in progress
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI、cs.LG
机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) ; South China University of Technology(华南理工大学)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI