arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-02-26 至 2026-02-26 共收录 204 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 19 篇

2511.12033 2026-02-26 cs.LG cs.AI 79%

EARL: Entropy-Aware RL Alignment of LLMs for Reliable RTL Code Generation

EARL: 用于可靠RTL代码生成的熵感知强化学习对齐大语言模型

Jiahe Shi, Zhengqi Gao, Ching-Yun Ko, Duane Boning

机构 * MIT(麻省理工学院) IBM(国际商业机器公司)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 EARL通过熵感知强化学习提升Verilog代码生成的可靠性与准确性

Comments Accepted to DAC 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09256 2026-02-26 cs.CV 78%

Hallucination Filtering in Radiology Vision-Language Models Using Discrete Semantic Entropy

在放射学视觉-语言模型中使用离散语义熵过滤幻觉

Patrick Wienholt, Sophie Caselitz, Robert Siepmann, Philipp Bruners, Keno Bressem, Christiane Kuhl, Jakob Nikolas Kather, Sven Nebelung, Daniel Truhn

机构 * Lab for Artificial Intelligence in Medicine, Department of Diagnostic and Interventional Radiology, University Hospital RWTH Aachen(医学人工智能实验室,诊断与介入放射学部,RWTH亚琛大学医院) Department of Diagnostic and Interventional Radiology, University Hospital RWTH Aachen(诊断与介入放射学部,RWTH亚琛大学医院) Department of Diagnostic and Interventional Radiology, Technical University of Munich, School of Medicine and Health, Klinikum rechts der Isar, TUM University Hospital(诊断与介入放射学部,慕尼黑技术大学,医学院与健康学院,Klinikum rechts der Isar,TUM大学医院) Department of Cardiovascular Radiology and Nuclear Medicine, Technical University of Munich, School of Medicine and Health, German Heart Center, TUM University Hospital(心血管放射学与核医学部,慕尼黑技术大学,医学院与健康学院,德国心脏中心,TUM大学医院) Else Kroener Fresenius Center for Digital Health, Faculty of Medicine and University Hospital Carl Gustav Carus, TUD Dresden University of Technology(数字健康中心,医学院与卡尔·古斯塔夫·卡尔斯大学医院,德累斯顿技术大学) Department of Medicine I, Faculty of Medicine and University Hospital Carl Gustav Carus, TUD Dresden University of Technology(第一医学部,医学院与卡尔·古斯塔夫·卡尔斯大学医院,德累斯顿技术大学) Pathology & Data Analytics, Leeds Institute of Medical Research at St James’s, University of Leeds(病理学与数据分析,圣詹姆斯医院医学研究所,利兹大学) Medical Oncology, National Center for Tumor Diseases (NCT), University Hospital Heidelberg(医学肿瘤学,国家肿瘤疾病中心(NCT),海德堡大学医院)

专题命中 知识编辑与模型理解 :language model(title,abstract)

AI总结 本研究通过离散语义熵过滤高熵问题,显著提升放射学视觉-语言模型的诊断准确性。

Comments Code is available: https://github.com/TruhnLab/VisionSemanticEntropy

Journal ref Eur Radiol (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19610 2026-02-26 cs.CV 78%

JailBound: Jailbreaking Internal Safety Boundaries of Vision-Language Models

JailBound: 视觉语言模型内部安全边界的劫持

Jiaxin Song, Yixu Wang, Jie Li, Rui Yu, Yan Teng, Xingjun Ma, Yingchun Wang

机构 * Shanghai Jiao Tong University(上海交通大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Fudan University(复旦大学) Xuan Tong(宣通) NSFOCUS

专题命中 知识编辑与模型理解 :language model(title,abstract)

AI总结 JailBound通过在视觉语言模型的潜在空间中探索安全边界,提出了一种新的劫持框架,有效提升了白盒和黑盒攻击成功率,揭示了模型的安全风险。

Comments The Thirty-ninth Annual Conference on Neural Information Processing Systems (NeurIPS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.21442 2026-02-26 cs.LG cs.AI 73%

MINAR: Mechanistic Interpretability for Neural Algorithmic Reasoning

MINAR: 图神经网络中神经算法推理的机制可解释性

Jesse He, Helen Jenne, Max Vargas, Davis Brown, Gal Mishne, Yusu Wang, Henry Kvinge

机构 * Pacific Northwest National Laboratory, Richland, WA(太平洋西北国家实验室) Halıcıoğlu Data Science Institute, University of California, San Diego, San Diego, CA(哈利奇奥格鲁数据科学研究所,加州大学圣地亚哥分校) Department of Computer and Information Science, University of Pennsylvania, Pennsylvaina, PA(计算机与信息科学系,宾夕法尼亚大学) Department of Mathematics, University of Washington, Seattle, WA(数学系,华盛顿大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 MINAR是一种用于图神经网络中神经算法推理的机制可解释性工具,通过归因修补方法发现电路,揭示训练过程中的电路形成和剪枝机制。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.09886 2026-02-26 cs.CL cs.AI 73%

Probabilistic distances-based hallucination detection in LLMs with RAG

基于概率距离的LLM中幻觉检测方法(RAG)

Rodion Oblovatny, Alexandra Kuleshova, Konstantin Polev, Alexey Zaytsev

机构 * Markov Lab, Department of Mathematics(马尔可夫实验室,数学系) Computer Science, Saint-Petersburg University(计算机科学,圣彼得堡大学) AI Center, Skoltech(人工智能中心,斯克里普丘克技术学院) SB AI Lab(SB人工智能实验室) AI Center, Skoltech, Risk department, Sber(人工智能中心,斯克里普丘克技术学院,风险部门)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出了一种基于概率距离的LLM幻觉检测方法,利用提示标记与响应标记嵌入分布的距离来检测幻觉,具有高效性和可迁移性。

Comments Updated approach to constructing a hallucination detection score. Added results from experiments with the NLI task. The approach with trainable deep kernels has been removed, with a focus on the unsupervised approach

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.01085 2026-02-26 cs.CV cs.AI 70%

Learning What Matters: Prioritized Concept Learning via Relative Error-driven Sample Selection

学习关键要素:通过相对误差驱动的样本选择进行优先概念学习

Shivam Chandhok, Qian Yang, Oscar Manas, Kanishk Jain, Leonid Sigal, Aishwarya Agrawal

机构 * Mila - Québec AI Institute(魁北克人工智能研究所) University of British Columbia(不列颠哥伦比亚大学) Université de Montréal(蒙特利尔大学) Vector Institute for AI(人工智能矢量研究所) Canada CIFAR AI Chair(加拿大CIFAR人工智能主席)

专题命中 知识编辑与模型理解 :language model(abstract);instruction tuning(abstract);分类 cs.AI

AI总结 PROGRESS通过相对误差驱动的样本选择,实现高效视觉-语言模型的指令微调,减少数据和计算需求,提升学习效率和泛化能力。

Comments CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.21496 2026-02-26 cs.AI 70%

Beyond Refusal: Probing the Limits of Agentic Self-Correction for Semantic Sensitive Information

超越拒绝:探求代理自我校正对语义敏感信息的极限

Umid Suleymanov, Zaur Rajabov, Emil Mirzazada, Murat Kantarcioglu

机构 * Department of Computer Science, Virginia Tech(弗吉尼亚理工大学计算机科学系) School of IT and Engineering, ADA University(ADA大学信息科技与工程学院)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出SemSIEdit框架,通过代理代理迭代批评和重写敏感信息,减少泄露并平衡隐私与效用,揭示规模依赖的安全分歧和推理悖论。

Comments Under Review

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.21864 2026-02-26 cs.CV cs.AI cs.CL cs.GR 62%

DynamicGTR: Leveraging Graph Topology Representation Preferences to Boost VLM Capabilities on Graph QAs

DynamicGTR: 利用图拓扑表示偏好提升视觉语言模型在图问答中的能力

Yanbin Wei, Jiangyue Yan, Chun Kang, Yang Chen, Hua Liu, James Kwok, Yu Zhang

机构 * Southern University of Science and Technology(南方科技大学) Hong Kong University of Science and Technology(香港理工大学) Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳)) Beihang University(北京航空航天大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI

AI总结 DynamicGTR通过动态选择最优图拓扑表示,提升视觉语言模型在图问答中的零样本性能及跨任务迁移能力。

Comments CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.21377 2026-02-26 cs.CL 57%

Beyond Subtokens: A Rich Character Embedding for Low-resource and Morphologically Complex Languages

超越子词:一种丰富的字符嵌入用于低资源和形态复杂语言

Felix Schneider, Maria Gogolev, Sven Sickert, Joachim Denzler

机构 * Computer Vision Group(计算机视觉组) Friedrich Schiller University(弗里德里希-席勒大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

AI总结 本文提出了一种基于字符的丰富嵌入方法,用于提升低资源和形态复杂语言的自然语言处理性能。

Comments 12 content pages, 2 figures, 8 tables, one example textbox

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24072 2026-02-26 cs.CV cs.AI 57%

Uncovering Grounding IDs: How External Cues Shape Multimodal Binding

揭示地面ID:外部线索如何塑造多模态绑定

Hosein Hasani, Amirmohammad Izadi, Fatemeh Askari, Mobin Bagherian, Sadegh Mohammadian, Mohammad Izadi, Mahdieh Soleymani Baghshah

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI

AI总结 本文提出地面ID概念,揭示外部线索通过增强多模态绑定的注意力机制,提升跨模态定位精度并减少幻觉。

Comments Under review as a conference paper at ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.22120 2026-02-26 cs.CV 50%

GeoDiv: Framework For Measuring Geographical Diversity In Text-To-Image Models

GeoDiv:文本到图像模型中地理多样性测量的框架

Abhipsa Basu, Mohana Singh, Shashank Agnihotri, Margret Keuper, R. Venkatesh Babu

机构 * Vision and AI Lab, Indian Institute of Science, Bangalore, India(印度科学院视觉与人工智能实验室) Chair for Machine Learning, University of Mannheim(曼海姆大学机器学习主任) Max Planck Institute for Informatics, Saarland Informatics Campus, Saarbrücken, Germany(马克斯·普朗克信息研究所,萨尔兰州信息学院,萨尔布吕肯,德国)

专题命中 知识编辑与模型理解 :language model(abstract)

AI总结 GeoDiv通过评估文本到图像模型中的地理多样性,揭示模型在描绘地区时的偏见问题,提出可解释的度量框架以促进公平生成系统。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.21735 2026-02-26 cs.CV 50%

SigVLP: Sigmoid Volume-Language Pre-Training for Self-Supervised CT-Volume Adaptive Representation Learning

SigVLP:基于sigmoid体积-语言预训练的自监督CT体积自适应表示学习

Jiayi Wang, Hadrien Reynaud, Ibrahim Ethem Hamamci, Sezgin Er, Suprosanna Shit, Bjoern Menze, Bernhard Kainz

机构 * Friedrich-Alexander University Erlangen-Nürnberg(弗里德里希-亚历山大大学埃尔兰根-纽伦堡) Department of Quantitative Biomedicine, University of Zurich(苏黎世大学定量生物医学系) ETH AI Center, ETH Zurich(苏黎世联邦理工学院AI中心) International School of Medicine, Istanbul Medipol University(伊斯坦布尔梅迪波尔大学国际医学院) Department of Computing, Imperial College London(伦敦帝国学院计算机系)

专题命中 知识编辑与模型理解 :language model(abstract)

AI总结 SigVLP通过引入旋转位置嵌入和块级对齐方法,改进CT体积与文本的自监督表示学习,提升文本到体积对齐的精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.21405 2026-02-26 cs.CV 50%

Rectifying Geometry-Induced Similarity Distortions for Real-World Aerial-Ground Person Re-Identification

校正几何诱导的相似性失真以应对现实中的空天地人重识别

Kailash A. Hambarde, Hugo Proença

机构 * Instituto de Telecomunicações(电信研究所) Department of Computer Science, Universidade da Beira Interior(计算机科学系,贝拉内尔内大学)

专题命中 知识编辑与模型理解 :prompting(abstract)

AI总结 本文提出GIQT模块,通过校正几何诱导的相似性失真,提升空天地人重识别在极端几何条件下的鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 其他LLM 11 篇

2602.22145 2026-02-26 cs.HC cs.AI cs.CL 90%

When AI Writes, Whose Voice Remains? Quantifying Cultural Marker Erasure Across World English Varieties in Large Language Models

当AI写作时,谁的声音仍保留?量化世界英语变体中大型语言模型的文化标记消失

Satyam Kumar Navneet, Joydeep Chandra, Yong Zhang

机构 * Independent Researcher(独立研究者) BNRIST, Dept. of CST, Tsinghua University(BNRIST计算机科学与技术系,清华大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 研究揭示大型语言模型在处理不同英语变体时对文化标记的系统性消除现象,并提出量化指标与改进方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.21522 2026-02-26 q-bio.NC cs.AI cs.CL 88%

One Brain, Omni Modalities: Towards Unified Non-Invasive Brain Decoding with Large Language Models

一个大脑,多模态:迈向统一非侵入式脑解码的大型语言模型

Changli Tang, Shurui Li, Junliang Wang, Qinfan Xiao, Zhonghao Zhai, Lei Bai, Yu Qiao, Bowen Zhou, Wen Wu, Yuanning Li, Chao Zhang

机构 * Tsinghua University(清华大学) Shanghai AI Laboratory(上海人工智能实验室) ShanghaiTech University(上海科技大学)

专题命中 其他LLM :language model(title,abstract);large language model(title);LLM(abstract);分类 cs.CL、cs.AI

AI总结 NOBEL通过统一EEG/MEG与fMRI信号,利用大型语言模型实现多模态非侵入式脑解码,提升解码准确性和对视觉语义的解读能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25184 2026-02-26 cs.CL cs.AI cs.GT 84%

Incentive-Aligned Multi-Source LLM Summaries

对齐激励的多源大语言模型摘要

Yanchen Jiang, Zhe Feng, Aranyak Mehta

机构 * Harvard University(哈佛大学) Google Research(谷歌研究)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出TTS框架,通过激励对齐提升多源摘要的事实准确性与稳健性,同时保持流畅性。

Comments Accepted at ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.22066 2026-02-26 cs.LG cs.AI 81%

DualWeaver: Synergistic Feature Weaving Surrogates for Multivariate Forecasting with Univariate Time Series Foundation Models

DualWeaver: 用于基于单变量时间序列基础模型的多变量预测的协同特征编织替代方案

Jinpeng Li, Zhongyi Pei, Huaze Xue, Bojian Zheng, Chen Wang, Jianmin Wang

机构 * School of Software, BNRist, Tsinghua University(软件学院,北京理工大学,清华大学) Tencent(腾讯)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI、cs.LG

AI总结 DualWeaver通过结构对称的替代序列提升多变量预测的准确性和稳定性,利用共享特征融合模块和无参数解码实现高效预测。

Comments 16 pages. Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.05085 2026-02-26 cs.CL cs.AI cs.IR 79%

Multi-Head RAG: Solving Multi-Aspect Problems with LLMs

多头RAG:利用LLMs解决多方面问题

Maciej Besta, Ales Kubicek, Robert Gerstenberger, Marcin Chrapek, Roman Niggli, Patrik Okanovic, Yi Zhu, Patrick Iff, Michal Podstawski, Lucas Weitzendorf, Mingyuan Chi, Joanna Gajda, Piotr Nyczyk, Jürgen Müller, Hubert Niewiadomski, Torsten Hoefler

机构 * Department of Computer Science, ETH Zurich(苏黎世联邦理工学院计算机科学系) IDEAS Research Institute(IDEAS研究 institute) NASK National Research Institute(国家研究 institute)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 多头RAG通过利用Transformer多头注意力机制提升多方面查询的检索准确性,实现更高的检索成功率和下游生成性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.22001 2026-02-26 cs.RO 78%

Are Foundation Models the Route to Full-Stack Transfer in Robotics?

基础模型是否是机器人领域全栈迁移的途径?

Freek Stulp, Samuel Bustamante, João Silvério, Alin Albu-Schäffer, Jeannette Bohg, Shuran Song

机构 * Institute of Robotics and Mechatronics, German Aerospace Center (DLR)(机器人与机电研究所,德国航空航天中心(DLR)) Stanford AI Lab, Stanford University(斯坦福大学人工智能实验室)

专题命中 其他LLM :foundation model(title,abstract)

AI总结 本文探讨基础模型在机器人全栈迁移中的作用,分析其对不同迁移层次的影响及面临的挑战。

Comments 12 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.21406 2026-02-26 cs.CV 78%

Exploring Vision-Language Models for Open-Vocabulary Zero-Shot Action Segmentation

探索用于开放词汇零样本动作分割的视觉-语言模型

Asim Unmesh, Kaki Ramesh, Mayank Patel, Rahul Jain, Karthik Ramani

机构 * Purdue University(普渡大学) Birla Institute of Technology and Science (BITS) Hyderabad(比拉理工学院和科学研究院(BITS)海得拉巴)

专题命中 其他LLM :language model(title,abstract)

AI总结 本文提出开放词汇零样本时间动作分割方法,利用视觉-语言模型的零样本能力,通过帧-动作嵌入相似性和相似性矩阵时间分割实现无需训练的分割任务,展示了其在结构化时间理解中的潜力。

Comments ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19430 2026-02-26 cs.CV 71%

TherA: Thermal-Aware Visual-Language Prompting for Controllable RGB-to-Thermal Infrared Translation

TherA: 用于可控RGB到热红外转换的热感知视觉-语言提示

Dong-Guw Lee, Tai Hyoung Rhee, Hyunsoo Jang, Young-Sik Shin, Ukcheol Shin, Ayoung Kim

机构 * Seoul National University(首尔国立大学) Kyungpook National University(庆北国立大学) KENTECH

专题命中 其他LLM :prompting(title)

AI总结 TherA通过热感知视觉-语言提示方法,实现了可控的RGB到热红外转换,提升了翻译性能和细粒度控制能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.21939 2026-02-26 cs.CY cs.AI 70%

Hidden Topics: Measuring Sensitive AI Beliefs with List Experiments

隐藏主题:通过列表实验测量敏感的AI信念

Maxim Chupilkin

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文通过列表实验揭示大型语言模型对监控、酷刑和核打击等敏感问题的潜在态度,验证了该方法在检测AI隐藏信念的有效性。

Comments 14 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.15929 2026-02-26 cs.CL 70%

Emergence of a phonological bias in ChatGPT

聊天机器人ChatGPT中元音偏见的出现

Juan Manuel Toro

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 ChatGPT表现出类似人类的辅音偏见,其在不同语言中均能识别单词,表明语言处理机制的相似性。

Comments 15 pages, 1 figure, corrected typo

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.21486 2026-02-26 cs.HC 50%

StoryComposerAI: Supporting Human-AI Story Co-Creation Through Decomposition and Linking

StoryComposerAI:通过分解与链接支持人机故事共创

Shuo Niu, Dylan Clements, Marina Margalit Nemanov, Hyungsin Kim

专题命中 其他LLM :prompting(abstract)

AI总结 StoryComposerAI通过分解与链接范式,提升人机共创故事的控制感与内容一致性。

Comments Extended Abstracts of the 2026 CHI Conference on Human Factors in Computing Systems (CHI EA '26)

详情

展开后加载摘要…

URL PDF HTML 收藏