arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12193 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12193 篇

2603.13768 2026-03-17 cs.SD cs.CL 79%

Causal Tracing of Audio-Text Fusion in Large Audio Language Models

大型音频语言模型中音频-文本融合的因果追溯

Wei-Chih Chen, Chien-yu Huang, Hung-yi Lee

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

AI总结 研究通过因果追溯分析大型音频语言模型在音频理解过程中音频特征与文本上下文的融合机制,揭示不同模型在层间和token级的融合策略及信息瓶颈。

Comments Submitted to Interspeech 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05748 2026-03-12 cs.LG 79%

Communication Enables Cooperation in LLM Agents: A Comparison with Curriculum-Based Approaches

通信使LLM代理协作:与基于课程的方法的比较

Hachem Madmoun, Salem Lahlou

机构 * MBZUAI(穆扎布伊人工智能研究所)

专题命中 其他LLM :LLM(title,abstract);分类 cs.LG

AI总结 本文通过比较直接通信与课程学习方法,发现通信在促进LLM代理合作中更有效,而课程学习可能因设计选择影响对齐目标。

Comments Published in EACL 2026 - Corrected cooperation rates for two-stage communication conditions (96.7% and 100.0%, previously reported as 48.3% and 50.0% due to a denominator bug in the evaluation code). All other results unchanged

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00969 2026-03-10 cs.AI cs.SY eess.SY 79%

Integrating a Causal Foundation Model into a Prescriptive Maintenance Framework for Optimising Production-Line OEE

将因果基础模型整合到指令性维护框架中以优化生产线OEE

Felix Saretzky, Lucas Andersen, Thomas Engel, Fazel Ansari

机构 * Department of Engineering University of Luxembourg(工程系卢森堡大学) Department of Computer Science University of Luxembourg(计算机科学系卢森堡大学) Chair of Production and Maintenance Management TU Wien(生产与维护管理系维也纳技术大学)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI

AI总结 本文提出基于因果机器学习的模型,通过模拟潜在修复方案优化生产线OEE,解决传统预测模型无法识别故障根本原因的问题。

Comments 9 pages, 3 images, 1 table, conference paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.07717 2026-03-10 cs.AI cs.GT cs.HC 79%

Rigidity in LLM Bandits with Implications for Human-AI Dyads

在LLM老虎机中的刚性及其对人机双元体的影响

Haomiaomiao Wang, Tomás E Ward, Lili Zhang

机构 * Insight Research Ireland Centre for Data Analytics, Ireland(爱尔兰洞察研究爱尔兰数据分析中心) School of Computing, Dublin City University, Ireland(都柏林城市大学计算机学院)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

AI总结 研究发现LLM在老虎机任务中表现出刚性决策策略,通过计算建模揭示了低学习率和高逆温度的机制,为理解人机交互中的决策偏见提供了新视角。

Comments 13 pages, 5 figures, AICS conference https://aicsconf.org/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05996 2026-03-09 cs.CL 79%

Track-SQL: Enhancing Generative Language Models with Dual-Extractive Modules for Schema and Context Tracking in Multi-turn Text-to-SQL

Track-SQL: 通过双提取模块增强生成语言模型以在多轮文本到SQL中进行模式和上下文跟踪

Bingfeng Chen, Shaobin Shi, Yongqi Luo, Boyan Xu, Ruichu Cai, Zhifeng Hao

机构 * School of Computer Science, Guangdong University of Technology(广东技术大学计算机科学学院) Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ)(广东人工智能与数字经济实验室) Peng Cheng Laboratory(鹏城实验室) College of Science, Shantou University(汕头大学理学院)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

AI总结 Track-SQL通过双提取模块提升生成语言模型在多轮文本到SQL任务中的模式和上下文跟踪能力,实现性能显著提升。

Comments Accepted at the 2025 Conference of the Nations of the Americas Chapter of the Association for Computational Linguistics (NAACL 2025), Long Paper, 19 pages

Journal ref Proceedings of the 2025 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers), pp. 10690-10708. Association for Computational Linguistics, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01653 2026-03-09 cs.CV cs.AI 79%

MAP: Mitigating Hallucinations in Large Vision-Language Models with Map-Level Attention Processing

MAP: 通过地图级注意力处理缓解大视觉-语言模型中的幻觉

Chenxi Li, Yichen Guo, Benfang Qian, Jinhao You, Kai Tang, Yaosong Du, Zonghao Zhang, Xiande Huang

机构 * DAIL Tech(DAIL科技)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

AI总结 MAP通过地图级注意力处理提升大视觉-语言模型的事实一致性与性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04514 2026-03-06 cs.AI 79%

Progressive Refinement Regulation for Accelerating Diffusion Language Model Decoding

逐步细化调节以加速扩散语言模型解码

Lipeng Wan, Jianhui Gu, Junjie Ma, Jianguo Huang, Shiguang Sun, Siyuan Li, Xuguang Lan

机构 * Department of Artificial Intelligence, Xi'an Jiaotong University(人工智能系,西安交通大学) College of Computing and Data Science, Nanyang Technological University(计算与数据科学学院,南洋理工大学) School of Computer Science and Technology, Harbin Institute of Technology(计算机科学与技术学院,哈尔滨工业大学)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

AI总结 逐步细化调节通过动态控制细化规则提升扩散语言模型解码效率并保持生成质量。

Comments 19 pages, 10 figures, Code available upon publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.02952 2026-03-04 q-bio.GN cs.LG q-bio.CB 79%

Sparse autoencoders reveal organized biological knowledge but minimal regulatory logic in single-cell foundation models: a comparative atlas of Geneformer and scGPT

稀疏自编码器揭示了组织化的生物知识,但在单细胞基础模型中包含最小的调控逻辑:Geneformer和scGPT的比较图谱

Ihor Kendiukhov

机构 * Department of Computer Science, University of Tübingen(图宾根大学计算机科学系)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.LG

AI总结 本研究通过稀疏自编码器分析单细胞基础模型Geneformer和scGPT,揭示其包含组织化的生物知识,但缺乏显著的因果调控逻辑。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18962 2026-03-04 cs.HC cs.AI cs.CY cs.IR cs.MA 79%

NeuroWise: A Multi-Agent LLM "Glass-Box" System for Practicing Double-Empathy Communication with Autistic Partners

NeuroWise:一种多智能体大语言模型"玻璃箱"系统,用于与自闭症伙伴练习双 empathy 交流

Albert Tang, Yifan Mo, Jie Li, Yue Su, Mengyuan Zhang, Sander L. Koole, Koen Hindriks, Jiahuan Pei

机构 * Marriotts Ridge High School(马里奥茨岭高中) Vrije Universiteit Amsterdam(阿姆斯特丹自由大学) Cake Researcher(蛋糕研究员)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

AI总结 NeuroWise通过多智能体大语言模型帮助神经典型用户改善与自闭症伙伴的双 empathy 交流,减少缺陷归因并提高沟通效率。

Comments Accepted to ACM CHI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14086 2026-03-04 cs.CR cs.AI 79%

Every Language Model Has a Forgery-Resistant Signature

每个语言模型都有抗伪造的签名

Matthew Finlayson, Xiang Ren, Swabha Swayamdipta

机构 * University of Southern California(南加州大学)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

AI总结 本研究提出了一种基于语言模型输出椭圆表面的抗伪造签名方法,通过几何约束验证模型输出来源,具有高安全性与实用性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.17848 2026-02-23 cs.CL 79%

On the scaling relationship between cloze probabilities and language model next-token prediction

关于闭合概率与语言模型下一个词预测之间的缩放关系

Cassandra L. Jacobs, Morgan Grobol

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

AI总结 本研究探讨了语言模型在闭合概率与下一个词预测之间的缩放关系,发现更大模型在语义上更符合人类响应,但对低层信息不敏感。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.16033 2026-02-19 cs.HC cs.AI 79%

Transforming GenAI Policy to Prompting Instruction: An RCT of Scalable Prompting Interventions in a CS1 Course

将生成式AI政策转化为提示指令:一项在CS1课程中可扩展的提示干预随机对照试验

Ruiwei Xiao, Runlong Ye, Xinying Hou, Jessica Wen, Harsh Kumar, Michael Liut, John Stamper

机构 * Carnegie Mellon University(卡内基梅隆大学) University of Toronto(多伦多大学) University of Michigan(密歇根大学) University of Toronto Mississauga(多伦多大学滑铁卢分校)

专题命中 其他LLM :prompting(title,abstract);分类 cs.AI

AI总结 本研究通过随机对照试验验证了基于ICAP框架的提示干预在提升学生提示技能和考试成绩方面的有效性,为生成式AI教育政策的改进提供了理论和实践指导。

Comments 11 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14968 2026-02-17 cs.RO cs.AI 79%

PhyScensis: Physics-Augmented LLM Agents for Complex Physical Scene Arrangement

PhyScensis: 基于物理的LLM代理用于复杂物理场景布置

Yian Wang, Han Yang, Minghao Guo, Xiaowen Qiu, Tsun-Hsuan Wang, Wojciech Matusik, Joshua B. Tenenbaum, Chuang Gan

机构 * UMass Amherst(马萨诸塞大学阿姆赫斯特分校) Genesis AI MIT(麻省理工学院) MIT-IBM Watson AI Lab(麻省理工-IBM沃森人工智能实验室)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

AI总结 PhyScensis通过物理引擎驱动的LLM代理框架,生成复杂物理场景布局,提升机器人操作中的场景复杂性和物理准确性。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.00686 2026-02-17 cs.CV cs.AI 79%

LVLM-COUNT: Enhancing the Counting Ability of Large Vision-Language Models

LVLM-COUNT: 提升大视觉-语言模型的计数能力

Muhammad Fetrat Qharabagh, Mohammadreza Ghofrani, Kimon Fountoulakis

机构 * University of Waterloo(滑铁卢大学)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

AI总结 LVLM-COUNT通过分而治之的方法提升大视觉-语言模型对大量对象的计数能力,有效解决计数任务中的重复计数问题。

Comments 38 pages, 24 Figures, 19 Tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11754 2026-02-13 cs.MA cs.AI cs.GT 79%

Cooperation Breakdown in LLM Agents Under Communication Delays

在通信延迟下LLM代理间的合作破裂

Keita Nishimoto, Kimitaka Asatani, Ichiro Sakata

机构 * The University of Tokyo(东京大学)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

AI总结 研究探讨了通信延迟对LLM多智能体系统中合作的影响,发现延迟增加会引发代理剥削行为,但过度延迟反而促进合作,揭示了低层因素对合作的重要性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.10716 2026-02-12 eess.AS cs.CL cs.SD 79%

RE-LLM: Refining Empathetic Speech-LLM Responses by Integrating Emotion Nuance

RE-LLM:通过整合情绪细微差别来完善同理心语音-LLM响应

Jing-Han Chen, Bo-Hao Su, Ya-Tse Wu, Chi-Chun Lee

专题命中 其他LLM :LLM(title,abstract);分类 cs.CL

AI总结 RE-LLM通过整合情绪细微差别,提升语音-LLM的共情响应生成能力,实验显示在多个数据集上显著提高了情感反应和探索评分。

Comments 5 pages, 1 figure, 2 tables. Accepted at IEEE ASRU 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.09489 2026-02-11 cs.AI 79%

Computing Conditional Shapley Values Using Tabular Foundation Models

利用表格基础模型计算条件Shapley值

Lars Henry Berge Olsen, Dennis Christensen

机构 * Norwegian Computing Center(挪威计算中心) Norwegian Defence Research Establishment (FFI)(挪威国防研究机构(FFI))

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI

AI总结 本文利用TabPFN模型高效计算条件Shapley值,在模拟和真实数据上表现优异,运行效率高。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.07891 2026-02-10 cs.CV cs.AI 79%

Scalable Adaptation of 3D Geometric Foundation Models via Weak Supervision from Internet Video

通过互联网视频弱监督实现3D几何基础模型的可扩展适应

Zihui Gao, Ke Liu, Donny Y. Chen, Duochao Shi, Guosheng Lin, Hao Chen, Chunhua Shen

机构 * Zhejiang University, State Key Lab of CAD \& CG Nanyang Technological University Independent Researcher Zhejiang University of Technology

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI

AI总结 SAGE通过互联网视频弱监督实现几何基础模型的可扩展适应,提升零样本泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.06909 2026-02-09 cs.LG 79%

Revisiting the Generic Transformer: Deconstructing a Strong Baseline for Time Series Foundation Models

重新审视通用变换器:解构一个强大的时间序列基础模型基准

Yunshi Wen, Wesley M. Gifford, Chandra Reddy, Lam M. Nguyen, Jayant Kalagnanam, Anak Agung Julius

机构 * Rensselaer Polytechnic Institute(拉特格斯理工学院) IBM Research(IBM研究院)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.LG

AI总结 本文重新审视通用变换器,通过简单训练协议证明其在时间序列零样本预测中的优越性能,并通过消融研究揭示了模型扩展和数据组成对性能的关键影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.06373 2026-02-09 cs.CL 79%

ReBeCA: Unveiling Interpretable Behavior Hierarchy behind the Iterative Self-Reflection of Language Models with Causal Analysis

ReBeCA: 通过因果分析揭示语言模型迭代自我反思背后的可解释行为层级

Tianqiang Yan, Sihan Shang, Yuheng Li, Song Qiu, Hao Peng, Wenjian Luo, Jue Xie, Lizhen Qu, Yuan Gao

机构 * Monash University(莫纳什大学) Harbin Institute of Technology(哈尔滨工业大学) Hainan University(海南大学) South China University of Technology(华南理工大学) The Chinese University of Hong Kong(香港中文大学)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

AI总结 ReBeCA通过因果分析揭示语言模型自我反思中的可解释行为层级,识别因果关系以提升自我反思效果。

Comments 17 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03478 2026-02-04 cs.AI 79%

When Routing Collapses: On the Degenerate Convergence of LLM Routers

当路由崩溃:关于LLM路由的退化收敛

Guannan Lai, Han-Jia Ye

机构 * School of Artificial Intelligence, Nanjing University, China(人工智能学院,南京大学,中国) National Key Laboratory for Novel Software Technology, Nanjing University, China(新型软件技术国家重点实验室,南京大学,中国)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

AI总结 本文提出EquiRouter,一种决策意识的路由器,通过直接学习模型排名来缓解LLM路由中的路由崩溃问题,有效降低计算和货币成本。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.22513 2026-02-04 cs.AI 79%

Why Self-Rewarding Works: Theoretical Guarantees for Iterative Alignment of Language Models

为何自我奖励有效:语言模型迭代对齐的理论保证

Shi Fu, Yingjie Wang, Shengchao Hu, Peng Wang, Dacheng Tao

机构 * Nanyang Technological University(南洋理工大学)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

AI总结 本文通过理论分析揭示了自我奖励语言模型在迭代对齐中的有效性,证明了其性能随迭代次数呈指数衰减,并为线性softmax模型提供了定制化的理论保证。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01240 2026-02-03 cs.CL 79%

Minimizing Mismatch Risk: A Prototype-Based Routing Framework for Zero-shot LLM-generated Text Detection

降低匹配风险:一种基于原型的路由框架用于零样本LLM生成文本检测

Ke Sun, Guangsheng Bao, Han Cui, Yue Zhang

机构 * School of Engineering, Westlake University, China(西交大学工程学院,中国)

专题命中 其他LLM :LLM(title,abstract);分类 cs.CL

AI总结 本文提出DetectRouter框架,通过两阶段训练学习文本检测器的亲和力,以解决零样本LLM生成文本检测中的匹配风险问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00420 2026-02-03 cs.CV cs.AI cs.CR 79%

Text is All You Need for Vision-Language Model Jailbreaking

文本就是视觉-语言模型劫持所需

Yihang Chen, Zhao Xu, Youyuan Jiang, Tianle Zheng, Cho-Jui Hsieh

机构 * University of California, Los Angeles(加州大学洛杉矶分校)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

AI总结 Text-DJ通过将文本转换为图像并诱导干扰,成功绕过视觉-语言模型的安全防护,揭示了OCR能力在面对分散多模态输入时的脆弱性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.21900 2026-02-03 cs.CV cs.AI cs.CY cs.MM 79%

TraceRouter: Robust Safety for Large Foundation Models via Path-Level Intervention

TraceRouter: 通过路径级干预实现大型基础模型的鲁棒安全性

Chuancheng Shi, Shangze Li, Wenjun Lu, Wenhua Wu, Cong Wang, Zifeng Cheng, Fei Shen, Tat-Seng Chua

机构 * The University of Sydney, Sydney, Australia(悉尼大学) Nanjing University of Science(南京理工大学) Nanjing University, Nanjing, China(南京大学) National University of Singapore, Singapore, Singapore(新加坡国立大学)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI

AI总结 TraceRouter通过路径级干预有效切断有害语义的因果传播,提升大型基础模型的对抗鲁棒性与实用性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.22345 2026-02-02 cs.LG 79%

Failing to Explore: Language Models on Interactive Tasks

无法探索:交互任务中的语言模型

Mahdi JafariRaviz, Keivan Rezaei, Arshia Soltani Moakhar, Zahra Sodagar, Yize Cheng, Soheil Feizi

专题命中 其他LLM :language model(title,abstract);分类 cs.LG

AI总结 本研究发现语言模型在交互任务中存在系统性探索不足问题,通过并行执行和历史总结等轻量干预措施提升了探索效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09852 2026-01-27 cs.CL 79%

Bears, all bears, and some bears. Language Constraints on Language Models' Inductive Inferences

熊、所有熊,以及一些熊。语言模型归纳推理中的语言约束

Sriram Padmanabhan, Siyuan Song, Kanishka Misra

机构 * The University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

AI总结 该研究探讨了视觉语言模型在区分不同归纳推理类型时的语言约束,通过实验发现模型与人类在行为上具有一致性,差异源于归纳约束而非表层形式。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17050 2026-01-27 cs.CV cs.AI 79%

Single-Pixel Vision-Language Model for Intrinsic Privacy-Preserving Behavioral Intelligence

单像素视觉-语言模型用于内在隐私保护的行为智能

Hongjun An, Yiliang Song, Jiawei Shao, Zhe Sun, Xuelong Li

机构 * Institute of Artificial Intelligence (TeleAI), China Telecom(人工智能研究院(TeleAI),中国电信)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

AI总结 单像素视觉-语言模型通过内在隐私保护机制,在隐私敏感环境中实现安全监控与行为智能的平衡。

Comments Initial Version, Pending Updates. We welcome any feedback and suggestions for improvement. Please feel free to contact us at an.hongjun@foxmail.com

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.15351 2026-01-23 astro-ph.IM cs.AI 79%

OmniSpectra: A Unified Foundation Model for Native Resolution Astronomical Spectra

OmniSpectra:一种用于原分辨率天体光谱的统一基础模型

Md Khairul Islam, Judy Fox

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI

AI总结 OmniSpectra是一种能够处理任意长度光谱的统一基础模型,通过自适应架构和多尺度学习,在天文学任务中展现卓越的迁移学习能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14522 2026-01-22 cs.LG 79%

On the Runway Cascade of Transformers for Language Modeling

关于Transformer的运行道级联用于语言建模

Hunjae Lee, Corey Clark

机构 * Department of Computer Science, Southern Methodist University, Dallas TX USA(计算机科学系,南方 Methodist 大学,德克萨斯州达拉斯)

专题命中 其他LLM :language model(title,abstract);分类 cs.LG

AI总结 本文提出运行道意识重 wiring方法,通过直接融入运行道上下文改善Transformer的语言建模能力,提升信息检索与外推性能。

详情

展开后加载摘要…

URL PDF HTML 收藏