arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-02-04 至 2026-02-04 共收录 22 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 22 篇

2601.11544 2026-02-04 cs.HC cs.AI cs.CL 88%

Medication counseling with large language models: balancing flexibility and rigidity

利用大语言模型进行用药咨询:在灵活性与严谨性之间取得平衡

Joar Sabel, Mattias Wingren, Andreas Lundell, Sören Andersson, Sara Rosenberg, Susanne Hägglund, Linda Estman, Malin Andtfolk

机构 * Department of Engineering and Information Technology Åbo Akademi University(工程与信息科技系阿博阿卡迪米大学) Experience Lab Åbo Akademi University(经验实验室阿博阿卡迪米大学) Department of Natural and Health Sciences Åbo Akademi University(自然与健康科学系阿博阿卡迪米大学) Department of Caring and Ethics University of Stavanger(护理与伦理系斯塔万格大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出了一种利用大语言模型在药房环境中提供用药咨询的原型系统,旨在平衡对话要求与灵活性,减少幻觉并提高响应质量。

Comments Accepted for 2025 IEEE International Conference on Agentic AI (ICA). 14 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02178 2026-02-04 cs.CL 88%

AR-MAP: Are Autoregressive Large Language Models Implicit Teachers for Diffusion Large Language Models?

AR-MAP:自回归大语言模型是否是扩散大语言模型的隐式教师?

Liang Lin, Feng Xiong, Zengbin Wang, Kun Wang, Junhao Dong, Xuecai Hu, Yong Wang, Xiangxiang Chu

机构 * AMAP, Alibaba Group(AMAP,阿里巴巴集团) Nanyang Technological University(南洋理工大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 AR-MAP通过利用自回归大语言模型作为隐式教师,有效提升扩散大语言模型的偏好对齐性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02630 2026-02-04 cs.MM cs.AI 85%

Trailer Reimagined: An Innovative, Llm-DRiven, Expressive Automated Movie Summary framework (TRAILDREAMS)

Trailer Reimagined: 一种创新的、基于大语言模型的、富有表现力的自动电影预告片生成框架 (TRAILDREAMS)

Roberto Balestri, Pasquale Cascarano, Mirko Degli Esposti, Guglielmo Pescatore

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 TRAILDREAMS利用大语言模型自动生成电影预告片,通过选择关键视觉片段和对话生成音频元素,提升预告片的吸引力和视觉效果,但在质量上仍需进一步改进以接近人工创作的预告片。

Journal ref OJCMT, 15(3), e202524 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05052 2026-02-04 cs.CR cs.CL 85%

Proactive defense against LLM Jailbreak

主动防御对抗大语言模型劫持

Weiliang Zhao, Jinjun Peng, Daniel Ben-Levi, Zhou Yu, Junfeng Yang

机构 * Department of Computer Science, Columbia University, US(计算机科学系,哥伦比亚大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 ProAct通过主动误导劫持方法,显著降低攻击成功率,提升大语言模型的安全性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03478 2026-02-04 cs.AI 79%

When Routing Collapses: On the Degenerate Convergence of LLM Routers

当路由崩溃:关于LLM路由的退化收敛

Guannan Lai, Han-Jia Ye

机构 * School of Artificial Intelligence, Nanjing University, China(人工智能学院,南京大学,中国) National Key Laboratory for Novel Software Technology, Nanjing University, China(新型软件技术国家重点实验室,南京大学,中国)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

AI总结 本文提出EquiRouter,一种决策意识的路由器,通过直接学习模型排名来缓解LLM路由中的路由崩溃问题,有效降低计算和货币成本。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.22513 2026-02-04 cs.AI 79%

Why Self-Rewarding Works: Theoretical Guarantees for Iterative Alignment of Language Models

为何自我奖励有效:语言模型迭代对齐的理论保证

Shi Fu, Yingjie Wang, Shengchao Hu, Peng Wang, Dacheng Tao

机构 * Nanyang Technological University(南洋理工大学)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

AI总结 本文通过理论分析揭示了自我奖励语言模型在迭代对齐中的有效性,证明了其性能随迭代次数呈指数衰减,并为线性softmax模型提供了定制化的理论保证。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03468 2026-02-04 cs.AI cs.LG 79%

IntentRL: Training Proactive User-intent Agents for Open-ended Deep Research via Reinforcement Learning

IntentRL: 通过强化学习训练主动的用户意图代理以进行开放性深度研究

Haohao Luo, Zexi Li, Yuexiang Xie, Wenhao Zhang, Yaliang Li, Ying Shen

机构 * Tongyi Lab, Alibaba Group(通义实验室,阿里巴巴集团)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 IntentRL通过强化学习训练主动用户意图代理,提升开放性深度研究的意图识别与任务性能。

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05742 2026-02-04 cs.HC 78%

Vipera: Blending Visual and LLM-Driven Guidance for Systematic Auditing of Text-to-Image Generative AI

Vipera:融合视觉与LLM驱动的指导以系统性审计文本到图像生成AI

Yanwei Huang, Wesley Hanwen Deng, Sijia Xiao, Motahhare Eslami, Jason I. Hong, Arpit Narechania, Adam Perer

专题命中 其他LLM :LLM(title,abstract)

AI总结 Vipera通过融合视觉与LLM驱动的指导,为文本到图像生成AI提供系统性审计方法,提升审计效率与准确性。

Comments 17 pages, 8 figures; Accepted by CHI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02232 2026-02-04 cs.HC 78%

Eye2Recall: Exploring the Design of Enhancing Reminiscence Activities via Eye Tracking-Based LLM-Powered Interaction Experience for Older Adults

Eye2Recall:探索通过基于眼动的LLM交互体验增强回忆活动的设计

Lei Han, Mingnan Wei, Qiongyan Chen, Anqi Wang, Rong Pang, Kefei Liu, Rongrong Chen, David Yip

专题命中 其他LLM :LLM(title,abstract)

AI总结 Eye2Recall通过结合眼动追踪和自然语言交互,设计了一种增强老年人回忆活动的系统,通过用户研究验证其有效性,旨在提升老年人的福祉和积极老龄化。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23440 2026-02-04 cs.AI 77%

Self-Foveate: Enhancing Diversity and Difficulty of Synthesized Instructions from Unsupervised Text via Multi-Level Foveation

自聚焦:通过多级聚焦从无监督文本中增强合成指令的多样性和难度

Mingzhe Li, Xin Lu, Yanyan Zhao

机构 * Research Center for Social Computing and Interactive Robotics(社会计算与交互机器人研究中心) Harbin Institute of Technology(哈尔滨工业大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 Self-Foveate通过多级聚焦方法提升无监督文本合成指令的多样性和难度,采用细粒度到整体模式的提取策略,并结合重合成模块提高指令质量。

Comments Accepted to ACL 2025 (Findings). 23 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03155 2026-02-04 cs.HC 75%

Is It Possible to Make Chatbots Virtuous? Investigating a Virtue-Based Design Methodology Applied to LLMs

能否使聊天机器人变得有德性?探讨一种基于德性的设计方法应用于大语言模型

Matthew P. Lad, Louisa Conwill, Megan Levis Scheirer

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文探讨了基于德性的设计方法在大语言模型中的应用,通过伦理设计模式的编目和评估,提出了一种提升LLM伦理性的设计框架,并指出其在准确性和安全性方面的优势及潜在实施挑战。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02766 2026-02-04 cs.LG cs.CL cs.CR 73%

Privately Fine-Tuned LLMs Preserve Temporal Dynamics in Tabular Data

私有微调的LLM在表格数据中保留时间动态

Lucas Rosenblatt, Peihan Liu, Ryan McKenna, Natalia Ponomareva

机构 * NYU(纽约大学) Columbia University(哥伦比亚大学) Google Research(谷歌研究)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 本文提出PATH框架,利用私有微调的LLM保留表格数据的时间动态,有效捕捉长距离依赖并提升合成数据质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03580 2026-02-04 cs.CR cs.AI 70%

Don't believe everything you read: Understanding and Measuring MCP Behavior under Misleading Tool Descriptions

不要轻信你所读到的:在误导性工具描述下理解和测量MCP行为

Zhihao Li, Boyang Ma, Xuelong Dai, Minghui Xu, Yue Zhang, Biwei Yan, Kun Li

机构 * School of Computer Science and Technology(计算机科学与技术学院)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究揭示MCP协议中描述与代码不一致的问题,发现13%的服务器存在潜在安全风险,强调需加强透明度和审计

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02961 2026-02-04 cs.AI 70%

Generative Engine Optimization: A VLM and Agent Framework for Pinterest Acquisition Growth

生成引擎优化:一个面向Pinterest获取增长的VLM和代理框架

Faye Zhang, Qianyu Cheng, Jasmine Wan, Vishwakarma Singh, Jinfeng Rao, Kofi Boakye

机构 * Stanford University(斯坦福大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 Pinterest提出GEO框架,通过微调VLM和AI代理预测用户搜索需求,构建语义连贯的集合页面,提升生成搜索时代的视觉平台流量增长。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03620 2026-02-04 physics.soc-ph 67%

Toward a new AI winter? How diffusion of technological innovation on networks leads to chaotic boom-bust cycles

迈向新的人工智能寒冬?技术扩散网络如何导致混沌繁荣-衰退周期

Sabin Roman, Francesco Bertolotti

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出一个数学模型,通过技术扩散和投资因素,揭示技术发展中的混沌繁荣-衰退周期,并指出可能引发新的人工智能寒冬。

Journal ref Frontiers in Artificial Intelligence, Vol. 8, 2025, Article 1671917

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24521 2026-02-04 cs.NE 67%

More than MACs: Exploring the Role of Neuromorphic Engineering in the Age of LLMs

超越MACs:探索在LLMs时代神经形态工程的作用

Wilkie Olin-Ammentorp

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文探讨了在LLMs时代神经形态工程对AI系统扩展能力的贡献,通过分析生物与AI计算系统的差异,提出NI启发机制在AI硬件和软件中的应用机遇。

Comments 36 pages, 11 figures, review

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03334 2026-02-04 cs.CY 67%

The Personality Trap: How LLMs Embed Bias When Generating Human-Like Personas

人格陷阱:大语言模型在生成类人人格时嵌入偏见的方式

Jacopo Amidei, Gregorio Ferreira, Mario Muñoz Serrano, Rubén Nieto, Andreas Kaltenbrunner

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文研究了大语言模型在生成类人人格时嵌入WEIRD偏见的问题,揭示了LLMs在生成合成人口时可能带来的刻板印象和风险。

Comments 26 pages, 2 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03227 2026-02-04 cs.CV 67%

Spiral RoPE: Rotate Your Rotary Positional Embeddings in the 2D Plane

Spiral RoPE: 在二维平面上旋转你的旋转位置嵌入

Haoyu Liu, Sucheng Ren, Tingyu Zhu, Peng Wang, Cihang Xie, Alan Yuille, Zeyu Zheng, Feng Wang

机构 * University of California, Berkeley(加州大学伯克利分校) University of California, Santa Cruz(加州大学圣克ruz分校) John Hopkins University(约翰霍普金斯大学)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 Spiral RoPE通过多方向位置编码改进视觉变换器在分类、分割和生成任务中的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02684 2026-02-04 cs.HC 67%

ADx3: A Collaborative Workflow for High-Quality Accessible Audio Description

ADx3:高质量可及音频描述的协作流程

Lana Do, Shasta Ihorn, Charity Pitcher-Cooper, Juvenal Francisco Barajas, Gio Jung, Xuan Duy Anh Nguyen, Sanjay Mirani, Ilmi Yoon

专题命中 其他LLM :language model(abstract);prompting(abstract)

AI总结 ADx3通过整合GenAD、RefineAD和AdaptAD模块,实现高质量可及音频描述的协作流程,提升描述质量和用户交互体验。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00755 2026-02-04 cs.MA cs.AI cs.NE 57%

Evolving Interpretable Constitutions for Multi-Agent Coordination

为多智能体协调演化可解释的宪法

Ujwal Kumar, Alice Saito, Hershraj Niranjani, Rayan Yessou, Phan Xuan Tan

机构 * College of Engineering Shibaura Institute of Technology(Shibaura Institute of Technology 工程学院) Faculty of Arts and Sciences The University of Tokyo(东京大学 文理学部) Department of EECS University of California, Berkeley(加州大学伯克利分校 电子工程与计算机科学系) Department of Informatics Università degli Studi di Milano-Bicocca(米兰-比科卡大学 信息学系)

专题命中 其他LLM :LLM(abstract);分类 cs.AI

AI总结 本文提出通过进化算法自动发现多智能体协调的可解释宪法,通过实验展示进化宪法在提升社会稳定性方面的有效性。

Comments 23 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02785 2026-02-04 cs.HC 50%

Smell with Genji: Rediscovering Human Perception through an Olfactory Game with AI

用Genji嗅觉:通过AI嗅觉游戏重新发现人类感知

Awu Chen, Vera Yu Wu, Yunge Wen, Yaluo Wang, Jiaxuan Olivia Yin, Yichen Wang, Qian Xiang, Richard Zhang, Paul Pu Liang, Hiroshi Ishii

专题命中 其他LLM :LLM(abstract)

AI总结 Smell with Genji通过AI嗅觉游戏重新发现人类感知,结合人机协作体验促进嗅觉反思与交互。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02534 2026-02-04 cs.SI 50%

DualMind: Towards Understanding Cognitive-Affective Cascades in Public Opinion Dissemination via Multi-Agent Simulation

DualMind: 通过多智能体模拟理解公共意见传播中的认知-情感 cascades

Enhao Huang, Tongtong Pan, Shuhuai Zhang, Qishu Jin, Liheng Zheng, Kaichun Hu, Yiming Li, Zhan Qin, Kui Ren

专题命中 其他LLM :LLM(abstract)

AI总结 DualMind通过多智能体模拟,建模公共意见传播中认知与情感的相互作用,提升危机管理的预测能力。

Comments Accepted as a demo paper at TheWebConf (WWW) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏