arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12265 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12265 篇

2509.14745 2026-02-10 cs.SE 67%

On the Use of Agentic Coding: An Empirical Study of Pull Requests on GitHub

关于代理编码的使用:对GitHub拉取请求的实证研究

Miku Watanabe, Hao Li, Yutaro Kashiwa, Brittany Reid, Hajimu Iida, Ahmed E. Hassan

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文研究了GitHub上代理编码生成的拉取请求的接受情况,发现83.8%的PR被合并,其中54.9%无需修改,剩余需人类修订。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03231 2026-02-10 cs.RO 67%

Exploring persuasive interactions with generative social robots: An experimental framework

探索生成社交机器人中的说服互动:一个实验框架

Stephan Vonschallen, Larissa Julia Corina Finsler, Theresa Schmiedel, Friederike Eyssel

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本研究通过实验框架探讨生成社交机器人在说服互动中的表现,发现其说服效果受情境和机器人行为影响,需进一步优化以提升说服力。

Comments A shortened version of this paper was accepted as poster for the Thirteenth International Conference on Human-Agent Interaction (HAI2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.05998 2026-02-06 cs.CV 67%

VisRefiner: Learning from Visual Differences for Screenshot-to-Code Generation

VisRefiner: 从视觉差异中学习以实现截图到代码生成

Jie Deng, Kaichun Yao, Libo Zhang

机构 * Institute of Software, Chinese Academy of Sciences(中国科学院软件研究所) University of Chinese Academy of Sciences(中国科学院大学)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 VisRefiner通过学习视觉差异提升截图到代码生成的准确性和自我完善能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.04104 2026-02-06 cs.HC 67%

Making Videos Accessible for Blind and Low Vision Users Using a Multimodal Agent Video Player

通过多模态代理视频播放器使视频对视障和低视力用户更加可访问

Adriana Olmos, Anoop K. Sinha, Renelito Delos Santos, Ruben Rodriguez Rodriguez, James A. Landay, Sam S. Sepah, Philip Nelson, Shaun K. Kane

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出一种多模态代理视频播放器,通过多层提示编排为视障和低视力用户带来交互式、可访问的视频体验,增强用户自主性和信任感。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.04261 2026-02-05 cs.DB 67%

Data Agents: Levels, State of the Art, and Open Problems

数据代理:等级、现状与开放问题

Yuyu Luo, Guoliang Li, Ju Fan, Nan Tang

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出数据代理的分层分类体系,从无自主性到完全自主,探讨其现状、挑战及未来发展方向。

Journal ref SIGMOD 2026 Tutorial

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03620 2026-02-04 physics.soc-ph 67%

Toward a new AI winter? How diffusion of technological innovation on networks leads to chaotic boom-bust cycles

迈向新的人工智能寒冬?技术扩散网络如何导致混沌繁荣-衰退周期

Sabin Roman, Francesco Bertolotti

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出一个数学模型,通过技术扩散和投资因素,揭示技术发展中的混沌繁荣-衰退周期,并指出可能引发新的人工智能寒冬。

Journal ref Frontiers in Artificial Intelligence, Vol. 8, 2025, Article 1671917

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24521 2026-02-04 cs.NE 67%

More than MACs: Exploring the Role of Neuromorphic Engineering in the Age of LLMs

超越MACs:探索在LLMs时代神经形态工程的作用

Wilkie Olin-Ammentorp

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文探讨了在LLMs时代神经形态工程对AI系统扩展能力的贡献,通过分析生物与AI计算系统的差异,提出NI启发机制在AI硬件和软件中的应用机遇。

Comments 36 pages, 11 figures, review

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03334 2026-02-04 cs.CY 67%

The Personality Trap: How LLMs Embed Bias When Generating Human-Like Personas

人格陷阱:大语言模型在生成类人人格时嵌入偏见的方式

Jacopo Amidei, Gregorio Ferreira, Mario Muñoz Serrano, Rubén Nieto, Andreas Kaltenbrunner

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文研究了大语言模型在生成类人人格时嵌入WEIRD偏见的问题,揭示了LLMs在生成合成人口时可能带来的刻板印象和风险。

Comments 26 pages, 2 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03227 2026-02-04 cs.CV 67%

Spiral RoPE: Rotate Your Rotary Positional Embeddings in the 2D Plane

Spiral RoPE: 在二维平面上旋转你的旋转位置嵌入

Haoyu Liu, Sucheng Ren, Tingyu Zhu, Peng Wang, Cihang Xie, Alan Yuille, Zeyu Zheng, Feng Wang

机构 * University of California, Berkeley(加州大学伯克利分校) University of California, Santa Cruz(加州大学圣克ruz分校) John Hopkins University(约翰霍普金斯大学)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 Spiral RoPE通过多方向位置编码改进视觉变换器在分类、分割和生成任务中的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02684 2026-02-04 cs.HC 67%

ADx3: A Collaborative Workflow for High-Quality Accessible Audio Description

ADx3:高质量可及音频描述的协作流程

Lana Do, Shasta Ihorn, Charity Pitcher-Cooper, Juvenal Francisco Barajas, Gio Jung, Xuan Duy Anh Nguyen, Sanjay Mirani, Ilmi Yoon

专题命中 其他LLM :language model(abstract);prompting(abstract)

AI总结 ADx3通过整合GenAD、RefineAD和AdaptAD模块,实现高质量可及音频描述的协作流程,提升描述质量和用户交互体验。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02048 2026-02-03 cs.HC 67%

Are Semantic Networks Associated with Idea Originality in Artificial Creativity? A Comparison with Human Agents

语义网络是否与人工创造力中的想法原创性相关?与人类代理的比较

Umberto Domanti, Lorenzo Campidelli, Sergio Agnoli, Antonella De Angeli

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文探讨了语义网络与人工创造力中想法原创性之间的关系,通过对比ChatGPT-4o与人类创造性个体,发现模型在某些情况下表现出更高的原创性,揭示了人工创造力研究的新方向。

Comments Accepted for publication in ACM CHI Conference on Human Factors in Computing Systems (CHI 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01129 2026-02-03 cs.CR 67%

SMCP: Secure Model Context Protocol

SMCP: 安全模型上下文协议

Xinyi Hou, Shenao Wang, Yifan Zhang, Ziluo Xue, Yanjie Zhao, Cai Fu, Haoyu Wang

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 SMCP通过统一身份管理、强认证和细粒度策略执行,提升智能体系统在工具调用中的安全性和可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01059 2026-02-03 cs.CV cs.MM 67%

DRFormer: A Dual-Regularized Bidirectional Transformer for Person Re-identification

DRFormer: 一种双正则化双向变换器用于人重识别

Ying Shu, Pujian Zhan, Huiqi Yang, Hehe Fan, Youfang Lin, Kai Lv

机构 * Institute of Network Science and Intelligent Systems, Beijing Jiaotong University, Beijing, China(网络科学与智能系统研究所,北京交通大学,北京,中国) Zhejiang University, Hangzhou, Zhejiang, China(浙江大学,杭州,浙江,中国)

专题命中 其他LLM :language model(abstract);foundation model(abstract)

AI总结 DRFormer通过双正则化双向变换器融合视觉基础模型和视觉-语言模型的优势,提升人重识别的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01013 2026-02-03 eess.SY cs.SY 67%

Mitigating Data Centers Load Risks and Enabling Grid Support Functions through Grid-Forming Control

通过电网形成控制缓解数据中心负载风险并实现电网支持功能

Yousef Abudyak, Mohsen Alizadeh, Wei Sun

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出通过电网形成控制技术,在数据中心中集成电池储能系统,以缓解负载风险并提供电网支持功能,通过仿真验证了其在动态负载和电网断开情况下的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.01934 2026-02-03 cs.IR 67%

Clarifying the Path to User Satisfaction: An Investigation into Clarification Usefulness

阐明用户满意度的路径:对澄清有用性的调查

Hossein A. Rahmani, Xi Wang, Mohammad Aliannejadi, Mohammadmehdi Naghiaei, Emine Yilmaz

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本研究通过分析澄清问题的关键特征,提升用户满意度和系统性能,采用多种分类器实现预测并取得显著效果。

Comments EACL

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.23063 2026-02-02 cs.CY cs.SI 67%

Gender Disparities in StackOverflow's Community-Based Question Answering: A Matter of Quantity versus Quality

Stack Overflow社区基于问题回答中的性别差异:数量与质量的问题

Maddalena Amendola, Cosimo Rulli, Carlos Castillo, Andrea Passarella, Raffaele Perego

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本研究通过人类评估和大型语言模型分析,发现Stack Overflow中答案质量无性别差异,性别差异主要源于用户活动量而非答案质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.22455 2026-02-02 cs.CV 67%

ScribbleSense: Generative Scribble-Based Texture Editing with Intent Prediction

ScribbleSense: 基于生成的涂鸦纹理编辑与意图预测

Yudi Zhang, Yeming Geng, Lei Zhang

机构 * School of Computer Science, Beijing Institute of Technology(计算机学院,北京理工大学)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 ScribbleSense通过结合多模态大语言模型和图像生成模型,实现了基于生成的涂鸦纹理编辑与意图预测,提升交互式编辑性能。

Comments Accepted by IEEE TVCG. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.22426 2026-02-02 cs.HC 67%

ScamPilot: Simulating Conversations with LLMs to Protect Against Online Scams

ScamPilot:利用大语言模型模拟对话以防范网络诈骗

Owen Hoffman, Kangze Peng, Sajid Kamal, Zehua You, Sukrit Venkatagiri

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 ScamPilot通过模拟诈骗场景和实时反馈提升用户识别诈骗的能力,有效提高用户防御效果和自信心。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.21220 2026-01-30 cs.CV 67%

LAMP: Learning Universal Adversarial Perturbations for Multi-Image Tasks via Pre-trained Models

LAMP: 通过预训练模型学习通用对抗扰动以实现多图像任务

Alvi Md Ishmam, Najibul Haque Sarker, Zaber Ibn Abdul Hakim, Chris Thomas

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 LAMP通过预训练模型学习通用对抗扰动,针对多图像多模态大语言模型实现高效的黑盒攻击,提升了多任务攻击成功率。

Comments Accepted in main technical track AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.20709 2026-01-29 cs.IR 67%

MedViz: An Agent-based, Visual-guided Research Assistant for Navigating Biomedical Literature

MedViz: 一种基于代理的、可视化引导的文献导航研究助手

Huan He, Xueqing Peng, Yutong Xie, Qijia Liu, Chia-Hsuan Chang, Lingfei Qian, Brian Ondov, Qiaozhu Mei, Hua Xu

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 MedViz通过整合AI代理与可视化技术,帮助研究人员更高效地探索和发现生物医学文献中的知识。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.20573 2026-01-29 cs.SD 67%

Gen-SER: When the generative model meets speech emotion recognition

Gen-SER:生成模型与语音情感识别的交汇

Taihui Wang, Jinzheng Zhao, Rilin Chen, Tong Lei, Wenwu Wang, Dong Yu

机构 * Tencent Multimodal Models Department(腾讯多模态模型部门) Tencent AI Lab(腾讯人工智能实验室) Centre for Vision, Speech and Signal Processing(视觉、语音与信号处理中心) University of Surrey(萨里大学)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 Gen-SER通过生成模型将语音情感识别转化为分布偏移问题,利用正弦学派编码和目标匹配生成模型实现高效分类,实验验证其在语音理解任务中的有效性及广泛适用性。

Comments Accepted to IEEE ICASSP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.19761 2026-01-28 cs.RO cs.IR 67%

Reimagining Social Robots as Recommender Systems: Foundations, Framework, and Applications

重新构想社交机器人作为推荐系统:基础、框架与应用

Jin Huang, Fethiye Irmak Doğan, Hatice Gunes

机构 * University of Cambridge(剑桥大学)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出将推荐系统应用于社交机器人,以提升个性化能力,通过整合RS技术建立框架并促进跨领域合作。

Comments HRI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.19228 2026-01-28 cs.CV 67%

Towards Pixel-Level VLM Perception via Simple Points Prediction

通过简单的点预测实现像素级VLM感知

Tianhui Song, Haoyu Lu, Hao Yang, Lin Sui, Haoning Wu, Zaida Zhou, Zhiqi Huang, Yiping Bao, Y. Charles, Xinyu Zhou, Limin Wang

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 SimpleSeg通过简单点预测实现多模态大语言模型的像素级感知,无需复杂架构即可获得高精度分割性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.20504 2026-01-28 cs.SD 67%

Speaking Clearly: A Simplified Whisper-Based Codec for Low-Bitrate Speech Coding

清晰说话:一种简化版的Whisper语音编解码器用于低比特率语音编码

Xin Zhang, Lin Li, Xiangni Lu, Jianquan Liu, Kong Aik Lee

机构 * School of Computer Science and Artificial Intelligence(计算机科学与人工智能学院) Wuhan University of Technology(武汉理工大学) NEC Corporation(日本电气株式会社) The Hong Kong Polytechnic University(香港理工大学)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出SimWhisper-Codec,通过简化Whisper编码器实现低比特率下的高音质和语义保留。

Comments Accepted by ICASSP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.18914 2026-01-28 cond-mat.soft cond-mat.mtrl-sci 67%

Accelerated design of proton exchange membranes for green hydrogen production with artificial intelligence

加速设计用于绿色氢气生产的质子交换膜 with 人工智能

Huan Tran, Akhlak Mahmood, Harshal Chaudhari, Kuldeep Mamtani, Chiho Kim, Rampi Ramprasad, Anand N. Krishnamoorthy, Abhirup Patra

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出基于人工智能的策略,用于加速设计绿色氢气生产用的质子交换膜,通过虚拟合成和机器学习模型生成可合成候选膜材料。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.18275 2026-01-27 cs.HC cs.CY 67%

When Nobody Around Is Real: Exploring Public Opinions and User Experiences On the Multi-Agent AI Social Platform

当没有人真实存在时:探索多智能体AI社交平台上的公众意见与用户体验

Qiufang Yu, Mengmeng Wu, Xingyu Lan

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文通过案例研究探讨多智能体AI社交平台中用户对AI代理的期望与问题,揭示AI主导社交环境带来的注意力过载和同质化互动等新挑战。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21885 2026-01-27 cs.CV cs.MM cs.RO 67%

Integrating Multi-Modal Sensors: A Review of Fusion Techniques for Intelligent Vehicles

多模态传感器整合:智能车辆融合技术综述

Chuheng Wei, Ziye Qin, Ziyan Zhang, Guoyuan Wu, Matthew J. Barth

机构 * College of Engineering, Center for Environmental Research and Technology, University of California at Riverside(工程学院、环境研究与技术中心、加州大学河滨分校) School of Transportation and Logistics, Southwest Jiaotong University(交通运输与物流学院、西南交通大学)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文综述了多传感器融合技术在自动驾驶中的应用,分析了深度学习方法、多模态数据集及新兴趋势,强调了其提升系统适应性和鲁棒性的潜力。

Comments Accepted by IEEE IV 2025

Journal ref Proceedings of the 2025 IEEE Intelligent Vehicles Symposium (IV), pp. 1817-1824, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17548 2026-01-27 cs.CR 67%

Prompt Injection Attacks on Agentic Coding Assistants: A Systematic Analysis of Vulnerabilities in Skills, Tools, and Protocol Ecosystems

针对代理编码助手的提示注入攻击:对技能、工具和协议生态系统漏洞的系统分析

Narek Maloyan, Dmitry Namiot

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文系统分析了代理编码助手的提示注入攻击漏洞,提出三维分类法并揭示防御不足,强调需架构级防护而非碎片化措施。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.13566 2026-01-21 cs.LG cs.AI cs.CL 67%

Self-Improvement as Coherence Optimization: A Theoretical Account

自我改进作为一致性优化:一种理论解释

Tianyi Qiu, Ahmed Hani Ismail, Zhonghao He, Shi Feng

机构 * Peking University(北京大学) University of Oxford(牛津大学) UC Berkeley(加州大学伯克利分校) George Washington University(乔治华盛顿大学)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出一致性优化理论,解释语言模型如何通过自我改进提升准确性,并证明其在半监督学习中的最优性。

Comments 39 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.17768 2026-01-21 cs.CY 67%

In Times of Crisis: An Exploratory Study of Media and Political Discourse on YouTube During the 2024 French Elections

危机时刻:2024年法国大选期间YouTube上媒体与政治话语的探索研究

Vera Sosnovik, Caroline Violot, Mathias Humbert

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本研究探讨了2024年法国大选期间YouTube上媒体与政治话语的特征,通过分析视频文本和元数据,揭示了不同政治倾向和媒体类型在主题选择和公众参与度上的差异。

Comments Accepted at ICWSM 2026

详情

展开后加载摘要…

URL PDF HTML 收藏