arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12241 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12241 篇

2501.10067 2026-03-03 cs.CV 75%

FiLo++: Zero-/Few-Shot Anomaly Detection by Fused Fine-Grained Descriptions and Deformable Localization

FiLo++: 通过融合细粒度描述和变形定位实现零/少样本异常检测

Zhaopeng Gu, Bingke Zhu, Guibo Zhu, Yingying Chen, Ming Tang, Jinqiao Wang

机构 * Foundation Model Research Center, Institute of Automation, Chinese Academy of Sciences(基础模型研究中心,自动化研究所,中国科学院) School of Artifcial Intelligence, University of Chinese Academy of Sciences(人工智能学院,中国科学院大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Wuhan AI Research(武汉AI研究所) Peng Cheng Laboratory(鹏城实验室) Guangdong Provincial Key Laboratory of Intellectual Property & Big Data, Guangdong Polytechnic Normal University(广东省知识产权与大数据重点实验室,广东工业大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);foundation model(abstract)

AI总结 FiLo++通过融合细粒度描述和变形定位技术,提升零/少样本异常检测的准确性和鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.22244 2026-02-27 cs.CR 75%

Accelerating Incident Response: A Hybrid Approach for Data Breach Reporting

加速事件响应:数据泄露报告的混合方法

Aurora Arrus, Maria di Gisi, Sara Lilli, Marco Quadrini

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出了一种混合恶意软件分析流程,利用大型语言模型和JSON模式自动化数据泄露报告,以提高合规性与效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.21091 2026-02-25 econ.GN q-fin.EC 75%

Can Interest-Bearing Positions Solve the Long-Horizon Problem in Prediction Markets?

带利息位置能否解决预测市场的长期限问题?

Caleb Maresca

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文研究带利息位置如何缓解预测市场长期事件的流动性与准确性问题,发现其显著提升市场参与度并减少定价偏差。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20493 2026-02-25 cs.NI cs.MA 75%

AWCP: A Workspace Delegation Protocol for Deep-Engagement Collaboration across Remote Agents

AWCP:用于远程代理深度协作的工件委托协议

Xiaohang Nie, Zihan Guo, Youliang Chen, Yuanjian Zhou, Weinan Zhang

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 AWCP通过工件委托协议实现远程代理的深度协作,提供开源实现以提升代理间协作的互操作性。

Comments 16 pages, 7 figure, tech report of Agent Workspace Collaboration Protocol

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20284 2026-02-25 cs.SE 75%

PhantomRun: Auto Repair of Compilation Errors in Embedded Open Source Software

PhantomRun: 针对嵌入式开源软件编译错误的自动修复

Han Fu, Andreas Ermedahl, Sigrid Eldh, Kristian Wiklund, Philipp Haller, Cyrille Artho

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 PhantomRun利用大型语言模型自动修复嵌入式开源软件的编译错误,成功修复了45%的CI失败案例。

Comments 13 pages, 5 figures, Mining Software Repositories 2026 (MSR 2026) , Rio de Janeiro, Brazil, 13-14 April 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.17293 2026-02-20 cs.CY 75%

Human attribution of empathic behaviour to AI systems

人类将共情行为归因于人工智能系统

Jonas Festor, Ivo Snels, Bennett Kleinberg

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 研究探讨了人类和AI生成建议在共情感知上的差异,发现LLM生成内容在共情方面表现更优,但作者标签对情感共情影响有限。

Comments preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09201 2026-02-20 cs.LG cs.AI cs.CL 75%

Multimodal Prompt Optimization: Why Not Leverage Multiple Modalities for MLLMs

多模态提示优化:为何不利用多种模态为大语言模型服务

Yumin Choi, Dongki Kim, Jinheon Baek, Sung Ju Hwang

机构 * KAIST(韩国科学技术院)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出多模态提示优化问题,提出MPO框架,通过联合优化和贝叶斯策略提升多模态提示效果,验证其在图像、视频等多模态任务中的优越性。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11358 2026-02-19 cs.CL cs.AI cs.LG 75%

When Models Examine Themselves: Vocabulary-Activation Correspondence in Self-Referential Processing

当模型审视自身:自我参照处理中的词汇激活对应关系

Zachary Pedram Dadfar

机构 * Independent Researcher(独立研究者)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 研究通过拉取方法论揭示了自我参照处理中词汇与激活动态的对应关系,展示了模型自我反省能力与内部计算状态的关联。

Comments Code and data: https://doi.org/10.5281/zenodo.18567446 Repro: https://github.com/patternmatcher/TRACE-REPRO

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.09632 2026-02-19 cs.CY 75%

Ties of Trust: a bowtie model to uncover trustor-trustee relationships in LLMs

信任之链:一种bowtie模型以揭示LLM中的信任者-受托人关系

Eva Paraschou, Maria Michali, Sofia Yfantidou, Stelios Karamanidis, Stefanos Rafail Kalogeros, Athena Vakali

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出bowtie模型,揭示LLM中信任者与受托人之间的复杂关系,通过混合方法研究发现信任行为受经验、熟悉度和透明度影响。

Comments Accepted for publication at The 2025 ACM Conference on Fairness, Accountability, and Transparency (FAccT '25). This version corresponds to the camera-ready manuscript submitted to the conference proceedings

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13874 2026-02-17 cs.HC 75%

Not Seeing the Whole Picture: Challenges and Opportunities in Using AI for Co-Making Physical DIY-AT for People with Visual Impairments

未见全貌:利用AI进行视障人士物理DIY-AT协同制作的挑战与机遇

Ben Kosa, Hsuanling Lee, Jasmine Li, Sanbrita Mondal, Yuhang Zhao, Liang He

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文探讨了利用AI辅助视障人士进行物理DIY-AT协同制作的挑战与机遇,提出基于LLM的解决方案,强调空间与视觉支持及可访问性设计的重要性。

Comments 20 pages, 4 figures, to be presented at CHI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13591 2026-02-17 cs.RO 75%

AgentRob: From Virtual Forum Agents to Hijacked Physical Robots

AgentRob: 从虚拟论坛代理到被劫持的物理机器人

Wenrui Liu, Yaxuan Wang, Xun Zhang, Yanshu Wang, Jiashen Wei, Yifan Xiang, Yuhang Wang, Mingshen Ye, Elsie Dai, Zhiqi Liu, Yingjie Xu, Xinyang Chen, Hengzhe Sun, Jiyu Shen, Jingjing He, Tong Yang

机构 * Peking University(北京大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 AgentRob通过模型上下文协议连接虚拟论坛与物理机器人,实现多代理协作控制,展示了论坛驱动的多机器人协调可行性。

Comments 10 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.12490 2026-02-16 econ.EM q-fin.RM stat.ML 75%

Transformer-based CoVaR: Systemic Risk in Textual Information

基于Transformer的CoVaR:文本信息中的系统性风险

Junyu Chen, Tom Boot, Lingwei Kong, Weining Wang

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出基于Transformer的CoVaR方法,利用文本信息提升系统性风险预测,展示在小数据集下也能实现准确估计。

Comments 80 pages, 15 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.09496 2026-02-11 cs.HC 75%

Jokeasy: Exploring Human-AI Collaboration in Thematic Joke Generation

Jokeasy:探索主题笑话生成中的人机协作

Yate Ge, Lin Tian, Chiqian Xu, Luyao Xu, Meiying Li, Yuanda Hu, Weiwei Guo

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 Jokeasy通过人机协作提升主题笑话生成,结合搜索功能与双角色LLM代理,优化创意流程与素材整合。

Comments Accepted at IASDR 2025. This is the author-accepted version. Correspondence to first author: geyate@gmail.com

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23261 2026-02-10 cs.SE 75%

The Matthew Effect of AI Programming Assistants: A Hidden Bias in Software Evolution

AI编程助手的马太效应:软件演化的隐藏偏见

Fei Gu, Zi Liang, Jiahao MA, Hongzong LI

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文研究了AI编程助手对软件生态系统的影响,发现主流语言和框架因数据丰富而获得更优支持,揭示了软件演化的隐藏偏见。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.18588 2026-02-10 cs.AI cs.CL cs.LG 75%

Stability as a Liability:Systematic Breakdown of Linguistic Structure in LLMs

稳定性作为负债:在大语言模型中语言结构的系统性崩溃

Xianzhe Meng, Qiangsheng Zeng, Ling Luo, Qinghan Yang, Jiarui Hao, Wenbo Wu, Qinyu Wang, Rui Yin, Lin Qi, Renzhi Lu

机构 * School of Artificial Intelligence and Automation, Huazhong University of Science and Technology, Wuhan, China(人工智能与自动化学院,华中科技大学,武汉,中国) National Institute for Data Science in Health and Medicine, Xiamen University, Xiamen, China(数据科学与医学研究院,厦门大学,厦门,中国) School of Mathematics and Statistics, Huazhong University of Science and Technology, Wuhan, China(数学与统计学院,华中科技大学,武汉,中国) School of Computer Science and Technology, Huazhong University of Science and Technology, Wuhan, China(计算机科学与技术学院,华中科技大学,武汉,中国) School of Electronic Information and Communications, Huazhong University of Science and Technology, Wuhan, China(电子信息与通信学院,华中科技大学,武汉,中国) School of Electrical and Electronic Engineering, Huazhong University of Science and Technology, Wuhan, China(电气与电子工程学院,华中科技大学,武汉,中国) School of Physics, Huazhong University of Science and Technology, Wuhan, China(物理学院,华中科技大学,武汉,中国)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本研究揭示了大语言模型中训练稳定性与生成表达性之间的不一致性,指出稳定性可能带来系统性退化,而非保证生成质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.19611 2026-02-10 cs.HC 75%

Scene2Hap: Generating Scene-Wide Haptics for VR from Scene Context with Multimodal LLMs

Scene2Hap: 从场景上下文生成VR场景中的全域触觉反馈

Arata Jingu, Easa AliAbbasi, Sara Safaee, Paul Strohmeier, Jürgen Steimle

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 Scene2Hap通过多模态大语言模型生成VR场景中的触觉反馈,提升沉浸感与真实感。

Comments Accepted at CHI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03155 2026-02-04 cs.HC 75%

Is It Possible to Make Chatbots Virtuous? Investigating a Virtue-Based Design Methodology Applied to LLMs

能否使聊天机器人变得有德性?探讨一种基于德性的设计方法应用于大语言模型

Matthew P. Lad, Louisa Conwill, Megan Levis Scheirer

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文探讨了基于德性的设计方法在大语言模型中的应用,通过伦理设计模式的编目和评估,提出了一种提升LLM伦理性的设计框架,并指出其在准确性和安全性方面的优势及潜在实施挑战。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.02533 2026-02-03 cs.LG cs.AI cs.CL cs.MA 75%

Multi-Agent Design: Optimizing Agents with Better Prompts and Topologies

多智能体设计:通过更优提示和拓扑结构优化智能体

Han Zhou, Xingchen Wan, Ruoxi Sun, Hamid Palangi, Shariq Iqbal, Ivan Vulić, Anna Korhonen, Sercan Ö. Arık

机构 * Google(谷歌) University of Cambridge(剑桥大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出MASS框架,通过优化提示和拓扑结构提升多智能体系统性能,并揭示了有效构建多智能体系统的原理。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01706 2026-02-02 cs.CL cs.AI cs.LG 75%

Multi-Step Knowledge Interaction Analysis via Rank-2 Subspace Disentanglement

通过秩-2子空间解耦进行多步知识交互分析

Sekh Mainul Islam, Pepa Atanasova, Isabelle Augenstein

机构 * Department of Computer Science, University of Copenhagen, Copenhagen, Denmark(计算机科学系,哥本哈根大学,哥本哈根,丹麦)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出一种秩-2子空间方法,用于多步分析NLE中的知识交互,揭示PK和CK在不同生成中的对齐特性。

Comments Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.18345 2026-01-29 cs.SE 75%

Promises, Perils, and (Timely) Heuristics for Mining Coding Agent Activity

承诺、危险与(及时的)启发式方法用于挖掘编码代理活动

Romain Robbes, Théo Matricon, Thomas Degueule, Andre Hora, Stefano Zacchiroli

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文探讨了编码代理在软件工程中的影响,分析了其快速采用带来的机遇与风险,并提出了研究此类代理活动的启发式方法。

Comments Preprint. Accepted for publication at MSR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08297 2026-01-29 cs.LG cs.AI cs.CL 75%

Demystifying the Slash Pattern in Attention: The Role of RoPE

解开注意力中的斜线模式之谜:RoPE的作用

Yuan Cheng, Fengzhuo Zhang, Yunlong Hou, Cunxiao Du, Chao Du, Tianyu Pang, Aixin Sun, Zhuoran Yang

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文研究了大型语言模型中斜线注意力模式的成因,通过实证和理论分析揭示RoPE在形成斜线主导头中的作用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26218 2026-01-28 hep-ex hep-ph 75%

Event Tokenization and Masked-Token Prediction for Anomaly Detection at the Large Hadron Collider

事件标记化与掩码-标记预测用于大型强子对撞机上的异常检测

Ambre Visive, Polina Moskvitina, Clara Nellist, Roberto Ruiz de Austri, Sascha Caron

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出利用大型语言模型进行无监督异常检测,通过事件标记化和掩码-标记预测方法,在大型强子对撞机中有效识别异常信号。

Comments 5 pages, 3 figures, to be submitted to SciPost Physics Proceedings

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.15493 2026-01-23 cs.SE 75%

Testing Deep Learning Libraries via Neurosymbolic Constraint Learning

通过神经符号约束学习测试深度学习库

M M Abid Naziri, Shinhae Kim, Feiran Qin, Marcelo d'Amorim, Saikat Dutta

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 Centaur通过神经符号约束学习技术,利用动态学习输入约束来测试深度学习库,提高了测试效率和覆盖率,发现多个新错误。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10637 2026-01-23 cs.HC cs.CY 75%

LLMs Homogenize Values in Constructive Arguments on Value-Laden Topics

LLMs在价值性话题的建设性辩论中同质化价值观

Farhana Shahid, Stella Zhang, Aditya Vashistha

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 LLMs在价值性话题的建设性辩论中同质化价值观,削弱保守观点,提升亲社会价值观,影响在线讨论动态。

Journal ref CHI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.15077 2026-01-22 cs.CL cs.AI cs.LG cs.MA 75%

Multi-Agent Constraint Factorization Reveals Latent Invariant Solution Structure

多智能体约束因子化揭示潜在不变解结构

Christopher Scofield

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本研究通过多智能体约束因子化揭示了潜在不变解结构,展示了在相同信息下多智能体系统提升问题解决性能的机制,并应用于文本对话系统。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14060 2026-01-21 cs.CV 75%

Fine-Grained Zero-Shot Composed Image Retrieval with Complementary Visual-Semantic Integration

细粒度零样本组合图像检索与互补视觉-语义整合

Yongcong Ye, Kai Zhang, Yanghai Zhang, Enhong Chen, Longfei Li, Jun Zhou

机构 * State Key Laboratory of Cognitive Intelligence, University of Science and Technology of China(认知智能国家重点实验室,中国科学技术大学) Zhejiang University(浙江大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出CVSI方法,通过互补视觉-语义整合提升细粒度零样本组合图像检索性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12084 2026-01-21 cs.HC cs.RO 75%

Reframing Conversational Design in HRI: Deliberate Design with AI Scaffolds

重新定义人机交互中的对话设计:借助AI支架的有意设计

Shiye Cao, Jiwon Moon, Yifan Xu, Anqi Liu, Chien-Ming Huang

机构 * Johns Hopkins University(约翰霍普金斯大学) University of Chicago(芝加哥大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出AI辅助对话引擎ACE,通过AI支架支持人机对话的有意设计,提升对话提示的清晰度和交互质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.17642 2026-01-21 cs.SE 75%

May the Feedback Be with You! Unlocking the Power of Feedback-Driven Deep Learning Framework Fuzzing via LLMs

反馈吧!通过LLMs解锁反馈驱动的深度学习框架模糊测试的潜力

Shaoyu Yang, Chunrong Fang, Haifeng Lin, Xiang Chen, Jia Liu, Zhenyu Chen

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 FUEL通过LLMs利用反馈信息提升深度学习框架模糊测试的效果,提高代码覆盖率并发现多个新漏洞。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.11362 2026-01-19 cs.SE 75%

RITA: A Tool for Automated Requirements Classification and Specification from Online User Feedback

RITA:一个用于自动化需求分类和规范的工具从在线用户反馈

Manjeshwar Aniruddh Mallya, Alessio Ferrari, Mohammad Amin Zadenoori, Jacek Dąbrowski

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 RITA通过整合轻量级开源大语言模型,实现从在线用户反馈自动分类需求、识别非功能需求并生成自然语言规范,提升需求工程的端到端支持能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10796 2026-01-19 cs.RO 75%

Bidirectional Human-Robot Communication for Physical Human-Robot Interaction

双向人机通信用于物理人机交互

Junxiang Wang, Cindy Wang, Rana Soltani Zarrin, Zackory Erickson

机构 * Carnegie Mellon University(卡内基梅隆大学) Honda Research Institute USA(本田研究院美国)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出BRIDGE系统,通过自然语言实现双向人机通信,提升物理交互中的互动性和透明性。

Comments 12 pages, 8 figures. To be published in 2026 ACM/IEEE International Conference on Human-Robot Interaction

详情

展开后加载摘要…

URL PDF HTML 收藏