arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-01-21 至 2026-01-21 共收录 35 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 35 篇

2601.12099 2026-01-21 cs.CL cs.AI 90%

Large language models struggle with ethnographic text annotation

大型语言模型在民族志文本标注中表现不佳

Leonardo S. Goodall, Dor Shilton, Daniel A. Mullins, Harvey Whitehouse

机构 * Calleva Research Centre Oxford Internet Institute University of Oxford(牛津大学奥克斯福德互联网研究所卡列瓦研究中心) Cohn Institute for the History and Philosophy of Science and Ideas Tel Aviv University(特拉维夫大学科恩研究所) Birkbeck College University of London(伦敦大学伯克贝克学院) Centre for the Study of Social Cohesion University of Oxford(牛津大学社会凝聚力研究所以及牛津大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 本研究发现大型语言模型在民族志文本标注任务中表现不佳,无法替代人类专家。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26538 2026-01-21 cs.SE 89%

Empirical and Sustainability Aspects of Software Engineering Research in the Era of Large Language Models: A Reflection

大语言模型时代软件工程研究的经验与可持续性方面:一种反思

David Williams, Max Hort, Maria Kechagia, Aldeida Aleti, Justyna Petke, Federica Sarro

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract)

AI总结 本文反思了大语言模型时代软件工程研究在基准测试严谨性、可重复性及可持续性方面的挑战,并提出改进建议。

Comments 5 pages, Camera Ready Accepted at ICSE-NIER 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12555 2026-01-21 cs.CL 88%

Evaluating Contextually Mediated Factual Recall in Multilingual Large Language Models

评估多语言大语言模型中的上下文中介事实回忆

Yihong Liu, Bingyu Xiong, Hinrich Schütze

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 研究多语言大语言模型在自然上下文中回忆事实的能力,发现上下文中介显著降低事实回忆准确性,大模型更鲁棒,真实名称影响不系统。

Comments preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13855 2026-01-21 cs.CL cs.AI 86%

Harnessing Consistency for Robust Test-Time LLM Ensemble

利用一致性提升鲁棒性测试时LLM集成

Zhichen Zeng, Qi Yu, Xiao Lin, Ruizhong Qiu, Xuying Ning, Tianxin Wei, Yuchen Yan, Jingrui He, Hanghang Tong

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 CoRE通过利用模型一致性提升LLM集成的鲁棒性,通过token和model级别的一致性改进集成性能。

Comments 18 pages, 15 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.11572 2026-01-21 cs.LG cs.AI 86%

Discrete Semantic States and Hamiltonian Dynamics in LLM Embedding Spaces

离散语义状态与哈密顿动力学在大语言模型嵌入空间中的应用

Timo Aukusti Laine

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本研究通过哈密顿动力学分析LLM嵌入空间的结构,揭示了语义状态的离散性及潜在的量子力学联系。

Comments 23 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.13600 2026-01-21 cs.AI 85%

Foundations of Global Consistency Checking with Noisy LLM Oracles

基于噪声LLM预言机的全局一致性基础

Paul He, Elke Kirschbaum, Shiva Kasiviswanathan

机构 * Nanyang Technological University(南洋理工大学) Amazon Web Services(亚马逊网络服务)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出一种基于LLM的自适应分治算法,用于高效检测和定位自然语言事实集合的全局不一致问题,具有低次多项式查询复杂度。

Comments Under Review

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.02370 2026-01-21 cs.CY cs.CL 85%

Variance-Aware LLM Annotation for Strategy Research: Sources, Diagnostics, and a Protocol for Reliable Measurement

考虑方差的LLM注释用于策略研究:来源、诊断和可靠测量的协议

Arnaldo Camuffo, Alfonso Gambardella, Saeid Kazemi, Jakub Malachowski, Abhinav Pandey

机构 * Bocconi University, ION Management Science Lab.(博科尼大学,ION管理科学实验室)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出了一种考虑方差的LLM注释协议,通过诊断五个方差来源并制定采样预算和聚合规则,提升策略研究的注释可靠性与可重复性。

Comments 41 pages for the main paper 53 pages for appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07772 2026-01-21 cs.AI 85%

An approach for systematic decomposition of complex llm tasks

一种复杂大语言模型任务的系统分解方法

Tianle Zhou, Jiakai Xu, Guanhong Liu, Jiaxiang Liu, Haonan Wang, Eugene Wu

机构 * Columbia University(哥伦比亚大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出ACONIC框架,通过形式复杂度度量系统分解任务,提升大语言模型在复杂任务中的可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.11585 2026-01-21 cs.CL 85%

Entropic Context Shaping: Information-Theoretic Filtering for Context-Aware LLM Agents

熵 context 形状:基于信息论的 context-aware LLM agent 的过滤方法

Hyunjun Kim

机构 * KAIST(韩国科学技术院)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出熵 context 形状方法,通过信息论衡量 context 的语用效用,优于词汇相似性方法,在多轮 context 选择任务中取得显著提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06266 2026-01-21 cs.SE 85%

Self-Admitted Technical Debt in LLM Software: An Empirical Comparison with ML and Non-ML Software

大语言模型软件中的自我承认技术债:与机器学习和非机器学习软件的实证比较

Niruthiha Selvanayagam, Taher A. Ghaleb, Manel Abdellatif

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract)

AI总结 本文通过实证研究比较了大语言模型、机器学习和非机器学习软件中的技术债情况,发现LLM系统积累技术债的速度与ML系统相似,但保持无债时间更长,并发现了三种新的技术债类型。

Comments Accepted to SANER 2026 (IEEE International Conference on Software Analysis, Evolution and Reengineering)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.11893 2026-01-21 cs.CR 85%

Taming Various Privilege Escalation in LLM-Based Agent Systems: A Mandatory Access Control Framework

驯服各种基于大语言模型的代理系统中的特权提升:一种强制访问控制框架

Zimo Ji, Daoyuan Wu, Wenyuan Jiang, Pingchuan Ma, Zongjie Li, Yudong Gao, Shuai Wang, Yingjiu Li

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract)

AI总结 本文提出SEAgent框架,通过强制访问控制解决LLM代理系统中的特权提升问题,有效阻止攻击并保持低误报率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14098 2026-01-21 eess.SY cs.SY 82%

A flexible language model-assisted electronic design automation framework

一种灵活的语言模型辅助电子设计自动化框架

Cristian Sestito, Panagiota Kontou, Pratibha Verma, Atish Dixit, Alexandros D. Keros, Michael O'Boyle, Christos-Savvas Bouganis, Themis Prodromakis

专题命中 其他LLM :language model(title,abstract);large language model(abstract)

AI总结 本文提出一种基于语言模型的灵活EDA框架,通过生成兼容商业工具的文件并优化设计,提升多领域电子设计的自动化能力。

Comments 17 pages, 5 figures, 1 Supplementary (12 pages, 13 figures, 1 table)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12317 2026-01-21 cs.LG cs.AI 81%

Explanova: Automatically Discover Data Insights in N \times M Table via XAI Combined LLM Workflow

Explanova:通过XAI结合LLM工作流自动在N×M表格中发现数据洞察

Yiming Huang

机构 * Institute for Clarity in Documentation(清晰文档研究所) Inria Paris-Rocquencourt(巴黎-罗克Quantin 国家信息与自动化研究所) Rajiv Gandhi University(拉吉夫·甘地大学) Tsinghua University(清华大学) Palmer Research Laboratories(帕尔默研究实验室)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI、cs.LG

AI总结 Explanova通过结合XAI和LLM工作流,在N×M表格中实现更经济的自动数据洞察发现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.21486 2026-01-21 cs.AI 79%

Hypothesis Generation via LLM-Automated Language Bias for ILP

通过LLM自动语言偏见进行假设生成

Yang Yang, Jiemin Wu, Yutao Yue

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

AI总结 本文提出通过LLM自动设计语言偏见,结合ILP求解器生成可解释的逻辑规则,提升假设生成的性能和鲁棒性。

Comments accepted by AAAI 2026 Bridge LMReasoning

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.16681 2026-01-21 eess.AS cs.CL cs.SD 79%

Emotional Dimension Control in Language Model-Based Text-to-Speech: Spanning a Broad Spectrum of Human Emotions

基于语言模型的文本到语音情感维度控制:跨越人类情感的广泛光谱

Kun Zhou, You Zhang, Dianwen Ng, Shengkui Zhao, Hao Wang, Bin Ma

机构 * Tongyi Lab, Alibaba Group, Singapore(阿里云实验室,阿里巴巴集团,新加坡) University of Rochester, United States of America(罗切斯特大学,美国)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

AI总结 本文提出基于语言模型的TTS框架,通过PAD三维情感维度实现更广泛的情感表达,提升语音的自然度和多样性。

Comments ICASSP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06371 2026-01-21 econ.EM stat.AP 78%

The Promise of Time-Series Foundation Models for Agricultural Forecasting: Evidence from Commodity Prices

时间序列基础模型在农业预测中的潜力:以商品价格为证据

Le Wang, Boyuan Zhang

专题命中 其他LLM :foundation model(title,abstract)

AI总结 本文发现现代时间序列基础模型在农业预测中优于传统方法和USDA预测,其中Time-MoE在小麦和玉米预测上表现最佳。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21057 2026-01-21 cs.CR cs.LG 77%

Soft Instruction De-escalation Defense

软指令缓解防御

Nils Philipp Walter, Chawin Sitawarin, Jamie Hayes, David Stutz, Ilia Shumailov

机构 * CISPA Helmholtz Center for Information Security(CISPA信息安全研究中心) Google DeepMind(谷歌DeepMind) AI Sequrity Company(AI安全公司)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 SIC通过迭代提示净化循环,有效防御LLM在智能体系统中受到提示注入攻击,尽管存在局限性,但提升了系统安全性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14060 2026-01-21 cs.CV 75%

Fine-Grained Zero-Shot Composed Image Retrieval with Complementary Visual-Semantic Integration

细粒度零样本组合图像检索与互补视觉-语义整合

Yongcong Ye, Kai Zhang, Yanghai Zhang, Enhong Chen, Longfei Li, Jun Zhou

机构 * State Key Laboratory of Cognitive Intelligence, University of Science and Technology of China(认知智能国家重点实验室,中国科学技术大学) Zhejiang University(浙江大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出CVSI方法,通过互补视觉-语义整合提升细粒度零样本组合图像检索性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12084 2026-01-21 cs.HC cs.RO 75%

Reframing Conversational Design in HRI: Deliberate Design with AI Scaffolds

重新定义人机交互中的对话设计:借助AI支架的有意设计

Shiye Cao, Jiwon Moon, Yifan Xu, Anqi Liu, Chien-Ming Huang

机构 * Johns Hopkins University(约翰霍普金斯大学) University of Chicago(芝加哥大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出AI辅助对话引擎ACE,通过AI支架支持人机对话的有意设计,提升对话提示的清晰度和交互质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.17642 2026-01-21 cs.SE 75%

May the Feedback Be with You! Unlocking the Power of Feedback-Driven Deep Learning Framework Fuzzing via LLMs

反馈吧!通过LLMs解锁反馈驱动的深度学习框架模糊测试的潜力

Shaoyu Yang, Chunrong Fang, Haifeng Lin, Xiang Chen, Jia Liu, Zhenyu Chen

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 FUEL通过LLMs利用反馈信息提升深度学习框架模糊测试的效果,提高代码覆盖率并发现多个新漏洞。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12703 2026-01-21 cs.LG 74%

Towards Spectroscopy: Susceptibility Clusters in Language Models

朝向光谱学:语言模型中的易感性聚类

Andrew Gordon, Garrett Baker, George Wang, William Snell, Stan van Wingerden, Daniel Murfet

专题命中 其他LLM :language model(title);分类 cs.LG

AI总结 该研究通过易感性分析揭示语言模型中token的聚类特性,发现510个可解释聚类,验证了与稀疏自编码器的结构相似性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.13558 2026-01-21 cs.AI cs.CL 73%

Leveraging ChatGPT and Other NLP Methods for Identifying Risk and Protective Behaviors in MSM: Social Media and Dating apps Text Analysis

利用ChatGPT及其他NLP方法识别MSM中的风险与保护行为:社交媒体和约会应用文本分析

Mehrab Beikzadeh, Chenglin Hong, Cory J Cascalheira, Callisto Boka, Majid Sarrafzadeh, Ian W Holloway

机构 * 1 Department of Computer Science, Henry Samueli School of Engineering, University of California, Los Angeles, Los Angeles, CA 2 Department of Social Welfare, Luskin School of Public Affairs, University of California, Los Angeles, Los Angeles, CA 3 Department of Counseling \& Educational Psychology, New Mexico State University, Las Cruces, NM

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本研究利用NLP方法分析社交媒体和约会应用文本,以预测MSM的性风险行为和饮酒情况,展示大语言模型在公共卫生干预中的应用价值。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12051 2026-01-21 cs.CV 71%

A Unified Masked Jigsaw Puzzle Framework for Vision and Language Models

面向视觉和语言模型的统一遮蔽拼图框架

Weixin Ye, Wei Wang, Yahui Liu, Yue Song, Bin Ren, Wei Bi, Rita Cucchiara, Nicu Sebe

机构 * Beijing Jiaotong University(北京交通大学) Kuaishou(快手) Hong Kong University of Science and Technology(香港科技大学) Caltech(加州理工学院) University of Trento(特伦特大学) University of Modena and Reggio Emilia(摩德纳和雷吉奥艾米利亚大学)

专题命中 其他LLM :language model(title)

AI总结 本文提出MJP框架,通过随机token洗牌和未知位置嵌入遮蔽,提升Transformer模型在视觉和语言任务中的鲁棒性与性能。

Comments 9 figures, 12 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.00256 2026-01-21 cs.NI cs.AI 70%

Large AI Model-Enabled Secure Communications in Low-Altitude Wireless Networks: Concepts, Perspectives and Case Study

大AI模型赋能的低空无线网络安全通信:概念、视角与案例研究

Chuang Zhang, Geng Sun, Yijing Lin, Weijie Yuan, Sinem Coleri, Dusit Niyato

机构 * College of Computer Science and Technology, Jilin University(吉林大学计算机科学与技术学院) Singapore University of Technology and Design(新加坡科技设计大学) College of Computer Science and Technology, Key Laboratory of Symbolic Computation and Knowledge Engineering of Ministry of Education, Jilin University(吉林大学计算机科学与技术学院、教育部符号计算与知识工程重点实验室) College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算机与数据科学学院)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出基于大AI模型的优化框架,利用大语言模型提升低空无线网络安全通信的性能。

Comments This paper has been accepted to IEEE Communications Magazine

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.13566 2026-01-21 cs.LG cs.AI cs.CL 67%

Self-Improvement as Coherence Optimization: A Theoretical Account

自我改进作为一致性优化:一种理论解释

Tianyi Qiu, Ahmed Hani Ismail, Zhonghao He, Shi Feng

机构 * Peking University(北京大学) University of Oxford(牛津大学) UC Berkeley(加州大学伯克利分校) George Washington University(乔治华盛顿大学)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出一致性优化理论,解释语言模型如何通过自我改进提升准确性,并证明其在半监督学习中的最优性。

Comments 39 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.17768 2026-01-21 cs.CY 67%

In Times of Crisis: An Exploratory Study of Media and Political Discourse on YouTube During the 2024 French Elections

危机时刻:2024年法国大选期间YouTube上媒体与政治话语的探索研究

Vera Sosnovik, Caroline Violot, Mathias Humbert

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本研究探讨了2024年法国大选期间YouTube上媒体与政治话语的特征,通过分析视频文本和元数据,揭示了不同政治倾向和媒体类型在主题选择和公众参与度上的差异。

Comments Accepted at ICWSM 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.11335 2026-01-21 cs.CY 67%

Deception and Manipulation in Generative AI

生成AI中的欺骗与操纵

Christian Tarsney

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文探讨生成AI中的欺骗与操纵问题,提出更严格的监管标准和防御措施以防止AI生成内容的误导性行为。

Journal ref Philosophical Studies, Volume 182 (2025), pp 1865-87

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12035 2026-01-21 cs.SI 67%

Effective and Unsupervised Social Event Detection and Evolution via RAG and Structural Entropy

基于RAG和结构熵的有效且无监督的社会事件检测与演化

Qitong Liu, Hao Peng, Zuchen Li, Xihang Meng, Ziyu Yang, Jiting Li, Li Sun, Philip S. Yu

专题命中 其他LLM :language model(abstract);foundation model(abstract)

AI总结 RagSEDE通过引入代表性和多样性驱动的采样策略、基于RAG的新范式以及结构信息理论,有效解决了社交媒体中社会事件检测与演化中的三大挑战。

Comments 12 pages, 7 figures, accepted for The Web Conference (WWW) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05620 2026-01-21 cs.LG 57%

Hyperparameter Transfer Enables Consistent Gains of Matrix-Preconditioned Optimizers Across Scales

超参数转移使矩阵预条件优化器在不同尺度上实现一致的增益

Shikai Qiu, Zixi Chen, Hoang Phan, Qi Lei, Andrew Gordon Wilson

机构 * New York University(纽约大学)

专题命中 其他LLM :language model(abstract);分类 cs.LG

AI总结 通过超参数转移,研究发现不同规模下矩阵预条件优化器在训练大模型时能实现显著加速,而错误缩放会导致加速效果消失。

Comments NeurIPS 2025. Code available at: https://github.com/charliezchen/scaling-matrix-preconditioning

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.13894 2026-01-21 cs.SE 50%

Multi-Location Software Model Completion

多位置软件模型补全

Alisa Welter, Christof Tinnes, Sven Apel

专题命中 其他LLM :LLM(abstract)

AI总结 本文提出NextFocus,一种基于全局嵌入的神经网络,首次实现多位置软件模型补全,通过预测多个位置的变化提升补全效果。

Comments Accepted at the 48th IEEE/ACM International Conference on Software Engineering (ICSE 2026) - Research Track

详情

展开后加载摘要…

URL PDF HTML 收藏