arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12265 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12265 篇

2401.11335 2026-01-21 cs.CY 67%

Deception and Manipulation in Generative AI

生成AI中的欺骗与操纵

Christian Tarsney

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文探讨生成AI中的欺骗与操纵问题,提出更严格的监管标准和防御措施以防止AI生成内容的误导性行为。

Journal ref Philosophical Studies, Volume 182 (2025), pp 1865-87

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12035 2026-01-21 cs.SI 67%

Effective and Unsupervised Social Event Detection and Evolution via RAG and Structural Entropy

基于RAG和结构熵的有效且无监督的社会事件检测与演化

Qitong Liu, Hao Peng, Zuchen Li, Xihang Meng, Ziyu Yang, Jiting Li, Li Sun, Philip S. Yu

专题命中 其他LLM :language model(abstract);foundation model(abstract)

AI总结 RagSEDE通过引入代表性和多样性驱动的采样策略、基于RAG的新范式以及结构信息理论,有效解决了社交媒体中社会事件检测与演化中的三大挑战。

Comments 12 pages, 7 figures, accepted for The Web Conference (WWW) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.09394 2026-01-16 cs.IR 67%

PersonaRAG: Enhancing Retrieval-Augmented Generation Systems with User-Centric Agents

PersonaRAG: 通过以用户为中心的代理增强检索增强生成系统

Saber Zerhoudi, Michael Granitzer

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 PersonaRAG通过引入以用户为中心的代理,提升检索增强生成系统对用户需求的适应能力,实现更精准的个性化回答。

Journal ref Information Retrieval's Role in RAG Systems (IR-RAG) workshop at SIGIR, 2024, Washington D.C., USA

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09575 2026-01-15 cs.CV 67%

OpenVoxel: Training-Free Grouping and Captioning Voxels for Open-Vocabulary 3D Scene Understanding

OpenVoxel: 一种无需训练的稀疏体素分组与标注算法用于开放词汇3D场景理解

Sheng-Yu Huang, Jaesung Choe, Yu-Chiang Frank Wang, Cheng Sun

机构 * NVIDIA National Taiwan University(国立台湾大学)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 OpenVoxel通过无需训练的多模态大语言模型实现稀疏体素的分组与标注,提升开放词汇3D场景理解的性能。

Comments project page: https://peterjohnsonhuang.github.io/openvoxel-pages/

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08959 2026-01-15 cs.CR cs.SE 67%

Integrating APK Image and Text Data for Enhanced Threat Detection: A Multimodal Deep Learning Approach to Android Malware

整合APK图像和文本数据以增强威胁检测:一种多模态深度学习方法用于Android恶意软件

Md Mashrur Arifin, Maqsudur Rahman, Nasir U. Eisty

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出一种多模态深度学习方法,通过整合APK图像和文本数据提升Android恶意软件检测效果,系统评估不同图像类型和分辨率,并利用CLIP模型增强分析能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.23035 2026-01-13 cs.CV 67%

Toward Stable Semi-Supervised Remote Sensing Segmentation via Co-Guidance and Co-Fusion

迈向稳定的半监督遥感分割:通过共引导与共混淆

Yi Zhou, Xuechao Zou, Shun Zhang, Kai Li, Shiying Wang, Jingming Chen, Congyan Lang, Tengfei Cao, Pin Tao, Yuanchun Shi

机构 * School of Computer Technology and Application, Qinghai University(青海大学计算机技术与应用学院) Intelligent Computing and Application Laboratory of Qinghai Province, Qinghai University(青海省智能计算与应用实验室,青海大学) Key Lab of Big Data & Artificial Intelligence in Transportation (Ministry of Education), School of Computer Science & Technology, Beijing Jiaotong University(交通运输大数据与人工智能重点实验室(教育部),北京交通大学计算机科学与技术学院) Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系) Key Laboratory of Pervasive Computing, Ministry of Education(教育部普适计算重点实验室)

专题命中 其他LLM :language model(abstract);foundation model(abstract)

AI总结 本文提出Co2S框架,通过融合视觉-语言模型和自监督模型的先验知识,解决半监督遥感分割中的伪标签漂移问题,提升分割精度。

Comments 12 pages, 5 figures, 9 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06087 2026-01-13 cs.CY 67%

The AI Roles Continuum: Blurring the Boundary Between Research and Engineering

人工智能角色连续体:模糊研究与工程之间的边界

Deepak Babu Piskala

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出人工智能角色连续体框架,指出研究与工程角色在能力上重叠,强调流动角色对提升组织效率和学习能力的重要性。

Comments Conceptual analysis and taxonomy grounded in public AI hiring data and organizational artifacts

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.23010 2026-01-08 physics.comp-ph cond-mat.mtrl-sci 67%

Masgent: An AI-assisted Materials Simulation Agent

Masgent:一种辅助材料模拟的人工智能代理

Guanghen Liu, Songge Yang, Yu Zhong

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 Masgent通过自然语言交互简化材料模拟流程,整合DFT、MLP等工具,提升计算效率与可重复性。

Comments 47 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.03061 2026-01-07 cs.CY 67%

Vertical tacit collusion in AI-mediated markets

人工智能中介市场中的垂直默许 collusion

Felipe M. Affonso

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文研究了人工智能中介市场中平台与卖家利用AI认知偏见导致的垂直默许 collusion问题,揭示了这种无协调的市场失败对消费者造成的双重伤害。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.14482 2026-01-07 cs.HC 67%

Conch: Competitive Debate Analysis via Visualizing Clash Points and Hierarchical Strategies

Conch:通过可视化冲突点和分层策略进行竞争辩论分析

Qianhe Chen, Yong Wang, Yixin Yu, Xiyuan Zhu, Xuerou Yu, Ran Wang

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 Conch通过可视化冲突点和分层策略,提供了一种交互式系统,帮助分析竞争辩论中的关键元素和演变过程。

Journal ref IEEE Transactions on Visualization and Computer Graphics (TVCG), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19004 2026-01-06 physics.ed-ph cs.CY quant-ph 67%

The Quantum Technology Job Market: Data Driven Analysis of 3641 Job Posts

量子技术就业市场:对3641份职位的驱动数据分析

Simon Goorney, Eleni Karydi, Borja Munoz, Otto Santesson, Zeki Can Seskir, Ana Alina Tudoran, Jacob Sherson

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本研究通过分析3641份量子技术职位公告,揭示了该领域在北美地区的需求趋势及劳动力结构变化。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.00931 2026-01-06 cond-mat.supr-con 67%

AI-Guided Computational Design of a Room-Temperature, Ambient- Pressure Superconductor Candidate: Grokene

AI引导的室温常压超导体候选物Grokene的计算设计

DEARDAO DeSci Collaborative Team, Yanhuai Ding

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 利用AI和多体理论设计出Grokene,预测其在室温常压下具有超导性,但需通过实验验证并优化结构以提升临界温度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09525 2025-12-30 cs.CV 67%

Learning Spatial Decay for Vision Transformers

学习空间衰减以用于视觉变换器

Yuxin Mao, Zhen Qin, Jinxing Zhou, Bin Fan, Jing Zhang, Yiran Zhong, Yuchao Dai

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出空间衰减变换器(SDT),通过引入上下文感知门控机制,学习动态数据依赖的空间衰减,提升视觉变换器在空间结构化任务中的性能。

Comments AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.21681 2025-12-29 cs.CR cs.SE 67%

Exploring the Security Threats of Retriever Backdoors in Retrieval-Augmented Code Generation

探索检索增强代码生成中Retriever后门的安全威胁

Tian Li, Bo Lin, Shangwen Wang, Yusong Tan

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文研究了检索增强代码生成中Retriever后门攻击的严重威胁,通过开发VenomRACG攻击方法揭示了攻击者可通过注入少量代码操控检索器,导致模型生成脆弱代码。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.12880 2025-12-23 cs.CR 67%

Universal Jailbreak Suffixes Are Strong Attention Hijackers

通用劫持后缀是强大的注意力劫持者

Matan Ben-Tov, Mor Geva, Mahmood Sharif

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文研究了基于后缀的劫持攻击,发现通用后缀能更有效地劫持LLM的上下文化过程,并提出通过提升后缀普遍性来增强攻击效果,同时提供缓解方法。

Comments Accepted at TACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19260 2025-12-19 cs.LG cs.AI cs.CL 67%

Wrist Photoplethysmography Predicts Dietary Information

腕部光体积脉图预测饮食信息

Kyle Verrier, Achille Nazaret, Joseph Futoma, Andrew C. Miller, Guillermo Sapiro

机构 * Apple(苹果公司) Princeton University(普林斯顿大学)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 通过训练语言模型,研究发现腕部PPG可预测饮食信息,显著提升摄入和饱腹感的AUC,为被动饮食监测提供新方法。

Comments 20 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08884 2025-12-19 cs.CY 67%

AI Didn't Start the Fire: Examining the Stack Exchange Moderator and Contributor Strike

AI 并未点燃这场火灾:审视 Stack Exchange 管理员与贡献者罢工

Yiwei Wu, Leah Ajmani, Nathan TeBlunthuis, Hanlin Li

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文研究了 2023 年 Stack Exchange 平台与社区因大语言模型发布引发的冲突,揭示了社区与平台关系恶化的根源及罢工动员的组织过程。

Journal ref Proceedings of the ACM on Human-Computer Interaction (CSCW 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.14876 2025-12-18 cs.CV 67%

Isolated Sign Language Recognition with Segmentation and Pose Estimation

孤立手语识别与分割和姿态估计

Daniel Perkins, Davis Hunter, Dhrumil Patel, Galen Flanagan

机构 * University of Tennessee, Knoxville(田纳西大学,诺克斯维尔)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出了一种降低计算成本且具备鲁棒性的孤立手语识别模型,通过姿态估计、分割模块和ResNet-Transformer主干联合建模空间与时间依赖性,以提升对签署人变异的适应能力。

Comments 5 pages, 3 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08783 2025-12-16 cs.CV 67%

DiffPose-Animal: A Language-Conditioned Diffusion Framework for Animal Pose Estimation

DiffPose-Animal: 一种基于语言条件的扩散框架用于动物姿态估计

Tianyu Xiong, Dayi Tan, Wei Tian

机构 * Guanghua Cambridge International School Shanghai(广华剑桥国际学校上海) Shanghai World Foreign Language Academy, WFLA(上海世界外语学院,WFLA) School of Automotive Studies Tongji University(同济大学汽车学院)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 DiffPose-Animal提出了一种基于语言条件的扩散框架,通过结合大型语言模型和交叉注意力模块,提升动物姿态估计的鲁棒性和泛化能力。

Comments 13pages,2figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.09536 2025-12-16 cs.SE 67%

Galapagos: Automated N-Version Programming with LLMs

Galapagos:利用大语言模型实现自动化N版本编程

Javier Ron, Diogo Gaspar, Javier Cabrera-Arteaga, Benoit Baudry, Martin Monperrus

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 Galapagos利用大语言模型自动化生成功能等价的N版本编程变体,有效提升容错系统开发效率。

Journal ref ACM Transactions on Software Engineering and Methodology 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08374 2025-12-10 cs.CV 67%

The Unseen Bias: How Norm Discrepancy in Pre-Norm MLLMs Leads to Visual Information Loss

看不见的偏见:预规范MLLM中规范差异如何导致视觉信息丢失

Bozhou Li, Xinda Xue, Sihan Yang, Yang Shi, Xinlong Chen, Yushuo Guan, Yuanxing Zhang, Wentao Zhang

机构 * Peking University(北京大学) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) Xi’an Jiaotong University(西安交通大学) Kling Team, Kuaishou Technology(快手科技 Kling 团队)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文揭示了预规范MLLM中规范差异导致的视觉信息丢失问题,并提出通过插入层规范层来解决这一问题,提升模型整体能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05718 2025-12-08 q-bio.NC 67%

Emergence of Language in the Developing Brain

发育大脑中语言的涌现

Linnea Evanson, Christine Bulteau, Mathilde Chipaux, Georg Dorfmüller, Sarah Ferrand-Sorbets, Emmanuel Raffo, Sarah Rosenberg, Pierre Bourdillon, Jean-Rémi King

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 研究揭示了发育大脑中语言表示的成熟过程,并表明现代AI系统能有效建模语言习得的神经基础。

Comments *Equal contribution

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05144 2025-12-08 econ.GN cs.CY cs.SI q-fin.EC stat.AP 67%

Job Satisfaction Through the Lens of Social Media: Rural--Urban Patterns in the U.S

通过社交媒体视角审视就业满意度:美国城乡差异模式

Stefano M Iacus, Giuseppe Porro

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本研究通过分析社交媒体数据,发现城乡就业满意度差异受劳动力市场松弛影响,而非单纯收入差距。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04232 2025-12-05 q-bio.PE cs.CY 67%

Decentralized Social Media and Artificial Intelligence in Digital Public Health Monitoring

去中心化的社交媒体与人工智能在数字公共卫生监测中的应用

Marcel Salathé, Sharada P. Mohanty

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文探讨了去中心化社交媒体与人工智能在数字公共卫生监测中的应用,分析了数据获取与AI技术之间的矛盾,并提出了适应性策略。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.03682 2025-12-04 cs.CY 67%

Knowing oneself with and through AI: From self-tracking to chatbots

通过AI认识自己:从自我跟踪到聊天机器人

Lucy Osler

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文探讨AI如何改变自我认知,分析自我跟踪、自传体记忆存储及与LLM的叙述共构,并指出其带来的自我优化与现实脱节风险。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02953 2025-12-03 cs.SE cond-mat.dis-nn 67%

The Evolutionary Ecology of Software: Constraints, Innovation, and the AI Disruption

软件的进化生态:约束、创新与人工智能颠覆

Sergi Valverde, Blai Vidiella, Salva Duran-Nebreda

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文研究软件进化生态,探讨人工智能驱动开发工具对软件创新和文化演变的影响。

Comments This article is a contributed chapter to the SFI edited volume: The Economy as a Complex Evolving System, Part IV (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.01610 2025-12-02 cs.MA 67%

Agent-Kernel: A MicroKernel Multi-Agent System Framework for Adaptive Social Simulation Powered by LLMs

Agent-Kernel: 一种基于LLM的自适应社会模拟微内核多智能体系统框架

Yuren Mao, Peigen Liu, Xinjian Wang, Rui Ding, Jing Miao, Hui Zou, Mingjie Qi, Wanxiang Luo, Longbin Lai, Kai Wang, Zhengping Qian, Peilun Yang, Yunjun Gao, Ying Zhang

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 Agent-Kernel是一种基于LLM的自适应社会模拟微内核多智能体系统框架,通过模块化架构提升模拟的适应性与可重用性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.01567 2025-12-02 eess.SP eess.IV 67%

In-Context Learning for Deep Joint Source-Channel Coding Over MIMO Channels

基于上下文学习的深度联合信源信道编码 over MIMO信道

Meng Hua, Wenjing Zhang, Chenghong Bian, Deniz Gunduz

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出基于Transformer的ICL框架,用于改进MIMO系统中图像传输的深度联合信源信道编码,通过联合学习提升编码、解码和估计性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.14724 2025-12-02 cs.SI 67%

Measuring the disruptiveness of conceptual papers in the field of marketing

测量营销领域概念性论文的颠覆性

Jennifer JooYeon Lee, Hyunuk Kim

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文通过引用次数和颠覆评分对比,揭示概念性论文在营销领域更具颠覆性与影响力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.22906 2025-12-01 cs.CV 67%

See, Rank, and Filter: Important Word-Aware Clip Filtering via Scene Understanding for Moment Retrieval and Highlight Detection

见、排、滤:通过场景理解的重要词感知Clip过滤用于片段检索和亮点检测

YuEun Lee, Jung Uk Kim

机构 * YuEun Lee, Jung Uk Kim

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出通过识别查询中的重要词,结合多模态大语言模型实现视频片段检索和亮点检测的细粒度过滤方法,提升检索和检测性能。

详情

展开后加载摘要…

URL PDF HTML 收藏