arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-01-27 至 2026-01-27 共收录 404 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 18 篇

2601.17764 2026-01-27 cs.CL cs.AI 73%

Cross-Lingual Probing and Community-Grounded Analysis of Gender Bias in Low-Resource Bengali

跨语言探测与社区导向分析低资源孟加拉语中的性别偏见

Md Asgor Hossain Reaj, Rajan Das Gupta, Jui Saha Pritha, Abdullah Al Noman, Abir Ahmed, Golam Md Mohiuddin, Tze Hui Liew

机构 * Department of Computer Science(计算机科学系) AIUB, Dhaka, Bangladesh Department of Information Technology(信息科技系) Wilmington University, Delaware, USA WUST, Virginia, USA Faculty of Information Science(信息科学系) Technology Multimedia University, Melaka, Malaysia(技术多媒体大学, Melaka, Malaysia) Centre for Intelligent Cloud Computing(智能云计算中心)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本研究通过跨语言探测和社区导向分析,揭示了孟加拉语中性别偏见的特征,强调需本地化方法和社区参与以减少偏见。

Comments Accepted in 2025 4th International Conference on Smart Cities, Automation & Intelligent Computing Systems (ICON-SONICS)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17397 2026-01-27 cs.CL 70%

CLM-Bench: Benchmarking and Analyzing Cross-lingual Misalignment of LLMs in Knowledge Editing

CLM-Bench: 评估和分析LLMs在知识编辑中的跨语言不对齐

Yucheng Hu, Wei Zhou, Juesi Xiao

机构 * Tianjin University, School of Future Technology(天津大学,未来技术学院) Tianjin University, College of Intelligence and Computing(天津大学,智能计算学院)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 CLM-Bench通过构建文化意识的基准,揭示了LLMs在知识编辑中的跨语言不对齐问题,挑战了现有跨语言转移方法的有效性。

Comments EACL MME workshop paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17168 2026-01-27 cs.AI cs.MA 70%

Interpreting Agentic Systems: Beyond Model Explanations to System-Level Accountability

解释代理系统:超越模型解释到系统级问责

Judy Zhu, Dhari Gandhi, Himanshu Joshi, Ahmad Rezaie Mianroodi, Sedef Akinli Kocak, Dhanesh Ramachandran

机构 * Vector Institute for Artificial Intelligence(向量人工智能研究所) University of Texas, Austin(德克萨斯大学奥斯汀分校) Dalhousie University(达尔豪斯大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文探讨了代理系统中可解释性技术的必要性,提出需设计专门方法以确保系统在目标形成、环境交互和结果评估等阶段的可追溯性和问责性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05840 2026-01-27 cs.LG 70%

Multimodal Trajectory Representation Learning for Travel Time Estimation

多模态轨迹表示学习用于旅行时间估计

Zhi Liu, Xuyuan Hu, Xiao Han, Zhehao Dai, Zhaolin Deng, Guojiang Shen, Xiangjie Kong

机构 * Zhejiang University of Technology(浙江工业大学) Zhejiang University of Technology College of Computer Science(浙江工业大学计算机科学学院)

专题命中 知识编辑与模型理解 :language model(abstract);pretraining(abstract);分类 cs.LG

AI总结 本文提出多模态动态轨迹整合框架,通过整合GPS、网格轨迹和道路网络约束,提升旅行时间估计的性能和泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.18240 2026-01-27 cs.CV 67%

V-Loop: Visual Logical Loop Verification for Hallucination Detection in Medical Visual Question Answering

V-Loop:用于医学视觉问答中幻觉检测的视觉逻辑循环验证

Mengyuan Jin, Zehui Liao, Yong Xia

机构 * Northwestern Polytechnical University(西北工业大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

AI总结 V-Loop通过双向推理和视觉逻辑循环验证,提升医学视觉问答中幻觉检测的准确性和效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17690 2026-01-27 cs.SD cs.AI cs.IR cs.LG eess.AS 62%

Segment Length Matters: A Study of Segment Lengths on Audio Fingerprinting Performance

片段长度至关重要:对音频指纹识别性能中片段长度的研究

Ziling Gong, Yunyan Ouyang, Iram Kamdar, Melody Ma, Hongjie Chen, Franck Dernoncourt, Ryan A. Rossi, Nesreen K. Ahmed

机构 * Data Science Institute Columbia University(数据科学研究所 哥伦比亚大学) Dolby Laboratories(杜比实验室) Adobe Research(Adobe研究) Cisco Research(思科研究)

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.AI、cs.LG

AI总结 本文研究了片段长度对音频指纹识别性能的影响,发现短片段长度(0.5秒)表现更优,并评估了大语言模型推荐最佳片段长度的能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13379 2026-01-27 cs.AI cs.CV 57%

The Art of Saying "Maybe": A Conformal Lens for Uncertainty Benchmarking in VLMs

也许的艺术:一种用于VLMs不确定性基准测试的共形透镜

Asif Azad, Mohammad Sadat Hossain, MD Sadik Hossain Shanto, M Saifur Rahman, Md Rizwan Parvez

机构 * Bangladesh University of Engineering and Technology(孟加拉工程与技术大学) Qatar Computing Research Institute(卡塔尔计算研究所)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI

AI总结 本文提出了一种用于VLMs不确定性基准测试的共形透镜,评估了18种最先进的VLMs在6个多元数据集上的表现,发现更大模型在不确定性量化方面表现更佳,而数学和推理任务则表现出较差的不确定性性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17986 2026-01-27 cs.LG 57%

Federated learning for unpaired multimodal data through a homogeneous transformer model

通过同质Transformer模型实现无配对多模态数据的联邦学习

Anders Eklund

机构 * Department of Biomedical Engineering, Linköping University, Sweden(_linköping大学生物医学工程系) Department of Computer and Information Science, Linköping University, Sweden(_linköping大学计算机与信息科学系) Center for Medical Image Science and Visualization (CMIV), Linköping University, Sweden(_linköping大学医学影像科学与可视化中心)

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.LG

AI总结 本文提出通过同质Transformer模型实现联邦学习,解决无配对多模态数据的训练问题,通过公共锚点对齐和子空间稳定化微调方法,在不传输私有数据的情况下实现全局模型统一表示。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17770 2026-01-27 eess.SP cs.AI 57%

Context-Aware Iterative Token Detection and Masked Transmission for Wireless Token Communication

具有上下文感知的迭代令牌检测与掩码传输用于无线令牌通信

Junyong Shin, Joohyuk Park, Jihong Park, Jinho Choi, Yo-Seb Jeon

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI

AI总结 本文提出了一种基于预训练MLM的无线令牌通信框架,通过上下文感知的迭代检测和掩码策略提升通信质量与速率适应性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.19113 2026-01-27 cs.CL 57%

Argument-Based Consistency in Toxicity Explanations of LLMs

基于论证的一致性:大语言模型毒性解释的评估

Ramaravind Kommiya Mothilal, Joanna Roy, Syed Ishtiaque Ahmed, Shion Guha

机构 * University of Toronto(多伦多大学) trail-ml

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL

AI总结 本研究提出基于论证的一致性标准,评估大语言模型对毒性的推理能力,发现其在复杂关系下的解释不一致。

Comments 29 pages, 7 figures, 9 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19442 2026-01-27 cs.AI 57%

Style2Code: A Style-Controllable Code Generation Framework with Dual-Modal Contrastive Representation Learning

Style2Code: 一种具有双模态对比学习的可控制代码生成框架

Dutao Zhang, Nicolas Rafael Arroyo Arias, YuLong He, Sergey Kovalchuk

机构 * ITMO University(ITMO大学) Saint Petersburg University(圣彼得堡大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI

AI总结 Style2Code通过双模态对比学习和条件解码实现风格可控的代码生成,提升风格控制的同时保持代码正确性。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 其他LLM 33 篇

2601.18512 2026-01-27 cs.CL 89%

Using Large Language Models to Construct Virtual Top Managers: A Method for Organizational Research

利用大语言模型构建虚拟高层管理者:组织研究的一种方法

Antonio Garzon-Vico, Krithika Sharon Komalapati, Arsalan Shahid, Jan Rosier

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

AI总结 本文提出利用大语言模型构建虚拟高层管理者,用于组织研究,通过模拟决策过程验证其在道德判断上的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.18105 2026-01-27 cs.CR cs.AI 89%

Mitigating the OWASP Top 10 For Large Language Models Applications using Intelligent Agents

利用智能代理缓解大型语言模型应用中的OWASP Top 10问题

Mohammad Fasha, Faisal Abul Rub, Nasim Matar, Bilal Sowan, Mohammad Al Khaldy

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 本文提出利用智能代理缓解大型语言模型应用中的OWASP Top 10安全漏洞,通过实时检测和应对提升模型安全性。

Comments 5 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.11494 2026-01-27 cs.AI 89%

Mutagenesis screen to map the functions of parameters of Large Language Models

参数功能映射的突变筛选:大型语言模型参数功能研究

Yue Hu, Gang Hu, Jixin Zheng, Patrick X. Zhao, Ruimeng Wang

机构 * Genetics Branch, NCI, NIH(国家卫生研究院癌症研究所遗传学分支) Beijing Normal University(北京师范大学) Snap Inc.(Snap公司) Longevity Biomedical Inc.(长寿生物医学公司)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);prompting(abstract);分类 cs.AI

AI总结 通过突变筛选方法研究大型语言模型参数功能,揭示参数与功能的复杂关系及潜在扩展方向。

Comments 10 pages, 6 figures, supplementary material available online

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17016 2026-01-27 cs.CY cs.AI 89%

Measuring Political Stance and Consistency in Large Language Models

测量大型语言模型中的政治立场与一致性

Salah Feras Alali, Mohammad Nashat Maasfeh, Mucahid Kutlu, Saban Kardas

机构 * Department of Computer Science and Engineering(计算机科学与工程系) Qatar University(卡塔尔大学) Gulf Studies Center(海湾研究中心)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);prompting(abstract);分类 cs.AI

AI总结 本研究评估了九个大型语言模型在24个政治敏感问题上的立场和一致性,发现模型立场受提示技术影响,且某些问题立场稳定不变。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.21184 2026-01-27 cs.LG cs.AI cs.CL 87%

Jailbreak-as-a-Service++: Unveiling Distributed AI-Driven Malicious Information Campaigns Powered by LLM Crowdsourcing

Jailbreak-as-a-Service++:揭示由LLM众包驱动的分布式恶意信息活动

Yu Yan, Sheng Sun, Mingfeng Li, Yunlong Song, Xingzhou Zhang, Linran Lu, Zhifei Zheng, Min Liu, Qi Li

机构 * Institute of Computing Technology, CAS(中国科学院计算技术研究所) University of Chinese Academy of Sciences(中国科学院大学) China People’s Public Security University(中国人民公安大学) Tongji University(同济大学) South China Normal University(华南师范大学) Tsing Hua University(清华大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 Jailbreak-as-a-Service++研究通过LLM众包技术揭示分布式恶意信息活动,提出PoisonSwarm方法提升恶意任务生成效率并提出生态系统级防御需求。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.18375 2026-01-27 cs.CL 85%

Hierarchical Text Classification with LLM-Refined Taxonomies

基于LLM优化的分层文本分类

Jonas Golde, Nicolaas Jedema, Ravi Krishnan, Phong Le

机构 * Humboldt Universität zu Berlin(柏林洪堡大学) Amazon(亚马逊) Meta School of Computer Science, University of St Andrews(圣安德鲁大学计算机学院)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出TaxMorph框架,利用LLM优化分类体系,提升分层文本分类性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.15979 2026-01-27 cs.SE cs.AI 85%

OLAF: Towards Robust LLM-Based Annotation Framework in Empirical Software Engineering

OLAF:迈向面向经验软件工程的鲁棒LLM注释框架

Mia Mohammad Imran, Tarannum Shaila Zaman

机构 * Missouri University of Science and Technology(密苏里科技大学) University of Maryland Baltimore County(马里兰大学巴尔的摩县分校)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出OLAF框架,旨在提升基于LLM的注释在经验软件工程中的透明性和可重复性。

Journal ref 3rd International Workshop on Methodological Issues with Empirical Studies in Software Engineering (WSESE) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17021 2026-01-27 q-fin.PM cs.LG cs.MA 85%

Regret-Driven Portfolios: LLM-Guided Smart Clustering for Optimal Allocation

后悔驱动的组合:LLM引导的智能聚类用于最优分配

Muhammad Abro, Hassan Jaleel

机构 * Lahore University of Management Sciences(拉合尔管理科学大学) Institute for Clarity in Documentation(文档清晰研究所) Inria Paris-Rocquencourt(巴黎-罗quant研究所) Rajiv Gandhi University(拉贾·甘地大学) Tsinghua University(清华大学) Palmer Research Laboratories(帕勒尔研究实验室)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文提出一种LLM引导的智能聚类方法,通过整合在线学习动态和市场情绪指标,构建高夏普比率的投资组合,实现优于传统基准的收益和风险控制。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17769 2026-01-27 cs.HC 85%

Reflexa: Uncovering How LLM-Supported Reflection Scaffolding Reshapes Creativity in Creative Coding

Reflexa:揭示LLM支持的反思支架如何重塑创意编码中的创造力

Anqi Wang, Zhengyi Li, Lan Luo, Xin Tong, Pan Hui

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract)

AI总结 Reflexa通过系统化的反思支架提升创意编码中的创造力,通过结构化反思模式增强可控性、探索广度和原创性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17577 2026-01-27 cs.HC cs.AI cs.CL 81%

Status Hierarchies in Language Models

语言模型中的地位层级

Emilio Barkett

机构 * Brigham Young University–Hawaii COLUMBIA UNIVERSITY(哥伦比亚大学)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.AI

AI总结 语言模型在多智能体环境中会因地位线索形成层级,高地位分配反而降低高能力模型的服从,揭示AI系统中的新兴社会行为。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09852 2026-01-27 cs.CL 79%

Bears, all bears, and some bears. Language Constraints on Language Models' Inductive Inferences

熊、所有熊,以及一些熊。语言模型归纳推理中的语言约束

Sriram Padmanabhan, Siyuan Song, Kanishka Misra

机构 * The University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

AI总结 该研究探讨了视觉语言模型在区分不同归纳推理类型时的语言约束,通过实验发现模型与人类在行为上具有一致性,差异源于归纳约束而非表层形式。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17050 2026-01-27 cs.CV cs.AI 79%

Single-Pixel Vision-Language Model for Intrinsic Privacy-Preserving Behavioral Intelligence

单像素视觉-语言模型用于内在隐私保护的行为智能

Hongjun An, Yiliang Song, Jiawei Shao, Zhe Sun, Xuelong Li

机构 * Institute of Artificial Intelligence (TeleAI), China Telecom(人工智能研究院(TeleAI),中国电信)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

AI总结 单像素视觉-语言模型通过内在隐私保护机制,在隐私敏感环境中实现安全监控与行为智能的平衡。

Comments Initial Version, Pending Updates. We welcome any feedback and suggestions for improvement. Please feel free to contact us at an.hongjun@foxmail.com

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.18255 2026-01-27 cs.LG cs.AI 79%

Beyond Retention: Orchestrating Structural Safety and Plasticity in Continual Learning for LLMs

超越保留:在连续学习中协调结构安全与可塑性

Fei Meng

机构 * Yangtze Delta Region Institute of Tsinghua University(清华大学 Yangtze Delta 区研究所)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出OSW方法,通过正交子空间唤醒协调连续学习中的结构安全与可塑性,有效保留脆弱任务能力并维持新任务学习。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.05344 2026-01-27 cs.CV 78%

Coding the Visual World: From Image to Simulation Using Vision Language Models

将视觉世界编码:通过视觉语言模型从图像到模拟

Sagi Eppel

专题命中 其他LLM :language model(title,abstract)

AI总结 本文研究了视觉语言模型通过Im2Sim方法从图像生成模拟的能力,发现其在复杂系统建模上表现优异,但对细节复制能力有限。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17824 2026-01-27 cs.HC cs.IR 78%

OwlerLite: Scope- and Freshness-Aware Web Retrieval for LLM Assistants

OwlerLite:面向LLM助手的范围和新鲜度感知网络检索

Saber Zerhoudi, Michael Dinzinger, Michael Granitzer, Jelena Mitrovic

专题命中 其他LLM :LLM(title);language model(abstract)

AI总结 OwlerLite通过用户定义的范围和数据新鲜度提升LLM助手的检索可控性和可信度。

Journal ref Proceedings of the Companion Proceedings of the ACM Web Conference 2026 (WWW Companion '26)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17676 2026-01-27 cs.HC 78%

GazeSummary: Exploring Gaze as an Implicit Prompt for Personalization in Text-based LLM Tasks

GazeSummary: 探索目光作为隐式提示在基于文本的LLM任务中的个性化应用

Jiexin Ding, Yizhuo Zhang, Xinyun Liu, Ke chen, Yuntao Wang, Shwetak Patel, Akshay Gadre

专题命中 其他LLM :LLM(title,abstract)

AI总结 本文研究了利用用户目光作为隐式提示,通过LLM生成个性化文本摘要的方法,并验证了其在真实阅读任务中的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04427 2026-01-27 cs.SE cs.AI 77%

Speed at the Cost of Quality: How Cursor AI Increases Short-Term Velocity and Long-Term Complexity in Open-Source Projects

以质量为代价:Cursor AI如何在开源项目中提升短期速度和长期复杂性

Hao He, Courtney Miller, Shyam Agarwal, Christian Kästner, Bogdan Vasilescu

机构 * Carnegie Mellon University(卡内基梅隆大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文研究Cursor AI对开源项目开发速度和质量的影响,发现其短期提升速度但长期增加代码复杂性,指出质量保障是早期采用者的主要瓶颈。

Journal ref 23rd International Conference on Mining Software Repositories (MSR '26), April 13--14, 2026, Rio de Janeiro, Brazil

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.12945 2026-01-27 cs.CL cs.CV 77%

LLMPopcorn: Exploring LLMs as Assistants for Popular Micro-video Generation

LLMPopcorn:探索大型语言模型作为流行微视频生成助手

Junchen Fu, Xuri Ge, Kaiwen Zheng, Alexandros Karatzoglou, Ioannis Arapakis, Xin Xin, Yongxin Ni, Joemon M. Jose

机构 * University of Glasgow(格拉斯哥大学) Shandong University(山东大学) Amazon(亚马逊) Telefónica Research(Telefónica研究院) National University of Singapore(新加坡国立大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 LLMPopcorn利用大型语言模型生成流行微视频,通过实验证明先进LLM可生成高流行度内容,并优化生成效果。

Comments Accepted by ICASSP2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.18005 2026-01-27 math.CO cs.LG 77%

Flow-based Extremal Mathematical Structure Discovery

基于流的极值数学结构发现

Gergely Bérczi, Baran Hashemi, Jonas Klüver

机构 * Aarhus University, Denmark(奥胡斯大学,丹麦) Max Planck Institute for Mathematics in the Sciences, Leipzig, Germany(马克斯·普朗克数学研究所,莱比锡,德国)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 FlowBoost通过闭环生成框架结合流匹配模型、奖励引导策略优化和随机局部搜索,有效发现极值几何结构,减少训练资源和时间,超越现有方法。

Comments 32 pages

详情

展开后加载摘要…

URL PDF HTML 收藏