arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12169 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12169 篇

2602.23234 2026-06-09 cs.IR cs.AI cs.LG 版本更新 90%

Scaling Search Relevance: Augmenting App Store Ranking with LLM-Generated Judgments

扩展搜索相关性:用LLM生成的判断增强应用商店排名

Evangelia Christakopoulou, Vivekkumar Patel, Hemanth Velaga, Sandip Gaikwad, Sean Suchter, Venkat Sundaranatha

机构 * Apple(苹果公司)

专题命中 其他LLM :LLM(title,title_cn);分类 cs.AI、cs.LG

AI总结 针对应用商店排名中专家文本相关性标签稀缺的问题,通过微调LLM生成数百万标签,结合行为相关性优化排序器,显著提升Pareto前沿和转化率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10078 2026-06-01 cs.CL cs.AI 90%

Human Psychometric Questionnaires Mischaracterize LLM Behavior

人类心理测量问卷误判LLM行为

Woojung Song, Dongmin Choi, Yoonah Park, Jongwook Han, Eun-Ju Lee, Yohan Jo

机构 * Graduate School of Data Science, Seoul National University(首尔国立大学数据科学研究生院) Department of Communication, Interdisciplinary Program in Artificial Intelligence, Seoul National University(首尔国立大学通信系人工智能交叉学科项目)

专题命中 其他LLM :LLM(title,title_cn);分类 cs.CL、cs.AI

AI总结 通过比较LLM在Likert问卷和生成概率上的价值与人格特征,发现问卷存在系统性偏差,提出基于生成概率的评估方法更准确。

Comments 38 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.30152 2026-05-29 cs.CL cs.AI cs.HC 90%

Do Proactive Agents Really Need an LLM to Decide When to Wake and What to Anchor?

主动型智能体真的需要LLM来决定何时唤醒和锚定什么吗?

Xiaoze Liu, Ruowang Zhang, Amir H. Abdi, Michel Galley, Zhikai Chen, Siheng Xiong, Xiaoqian Wang, Jing Gao

机构 * Purdue University(普渡大学) Microsoft(微软) Michigan State University(密歇根州立大学) Georgia Institute of Technology(佐治亚理工学院)

专题命中 其他LLM :LLM(title,title_cn);分类 cs.CL、cs.AI

AI总结 提出用时间图学习(TGL)模型替代LLM作为主动智能体的触发器,通过图更新而非文本处理用户活动,实现高效、低延迟的触发决策。

Comments 31 pages, 5 figures, 7 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.19782 2026-05-20 cs.AI cs.LG cs.SE 90%

Prior Knowledge or Search? A Study of LLM Agents in Hardware-Aware Code Optimization

先验知识还是搜索?LLM代理在硬件感知代码优化中的研究

Dmitry Redko, Albert Fazlyev, Konstantin Sozykin, Maria Ivanova, Evgeny Burnaev, Egor Shvetsov

机构 * Applied AI Institute(应用人工智能研究所) ITMO University(ITMO大学) AI Talent Hub(AI人才中心)

专题命中 其他LLM :LLM(title,title_cn);分类 cs.AI、cs.LG

AI总结 该研究探讨了在硬件感知代码优化中,LLM代理是依赖于先验知识还是搜索过程,通过三个受控实验发现LLM在纯黑盒优化中表现为贪婪优化器,在零样本内核生成中输入大小信息无明显影响,而在反馈循环内核优化中CUDA单调改进而TVM IR主动退化,表明LLM在代码优化任务中高度依赖预训练先验而非反馈或代理结构。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06147 2026-05-14 cs.LG cs.CL stat.ML 90%

LLM Flow Processes for Text-Conditioned Regression

用于文本条件回归的LLM流程

Felix Biggs, Samuel Willis

机构 * Secondmind Wayve

专题命中 其他LLM :LLM(title,title_cn);分类 cs.CL、cs.LG

AI总结 本文提出结合轻量级神经过程与边际LLM预测,以解决文本条件回归中的误差累积问题,提升预测校准和局部一致性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.05808 2026-04-16 cs.AI cs.LG 90%

Hierarchical Reinforcement Learning with Augmented Step-Level Transitions for LLM Agents

具有增强步骤级转换的分层强化学习用于大语言模型代理

Shuai Zhen, Yanhua Yu, Ruopei Guo, Nan Cheng, Yang Deng

机构 * Beijing University of Posts and Telecommunications(北京邮电大学) China Mobile Group Design Institute Co., Ltd(中国移动集团设计院有限公司) Singapore Management University(新加坡管理大学)

专题命中 其他LLM :LLM(title,summary_cn);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出STEP-HRL框架,通过仅基于单步转换而非完整交互历史实现步骤级学习,提升LLM代理在复杂任务中的性能和泛化能力,同时减少token使用。

Comments Accepted to ACL 2026 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.10268 2025-11-19 cs.SE cs.AI cs.LG 90%

Generating Streamlining Constraints with Large Language Models

Florentina Voboril, Vaidyanathan Peruvemba Ramaswamy, Stefan Szeider

机构 * Algorithms and Complexity Group TU Wien(算法与复杂性组维也纳技术大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI、cs.LG

Comments 23 page; deeper analysis of streamliners and statistics about benchmark instances added

Journal ref F. Voboril, V. P. Ramaswamy, S. Szeider, Generating Streamlining Constraints with Large Language Models, Journal of Artificial Intelligence Research, volume 84, pages 16:1-16:19, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.16891 2024-04-29 cs.CR cs.AI cs.CL cs.CY 90%

Attacks on Third-Party APIs of Large Language Models

Wanru Zhao, Vidit Khazanchi, Haodi Xing, Xuanli He, Qiongkai Xu, Nicholas Donald Lane

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

Comments ICLR 2024 Workshop on Secure and Trustworthy Large Language Models

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.31281 2026-08-19 cs.CL 版本更新 90%

Wind Turbine Maintenance Log Labelling Framework: LLM-Driven Data Correction and Enrichment via Semantic Extraction of Reliability Intelligence

风力涡轮机维护日志标注框架:基于LLM驱动的数据校正与语义提取的可靠性智能增强

Max Malyi, Jonathan Shek, Alasdair McDonald, Andre Biscaya

机构 * Institute for Energy Systems, School of Engineering, The University of Edinburgh(能源系统研究所,工程学院,爱丁堡大学) Nadara, Lisbon, Portugal(纳达拉,里斯本,葡萄牙)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 提出一种利用大语言模型自动标准化和结构化风力涡轮机维护日志的方法,通过纠正系统代码、提取故障模式与维护动作分类,将非结构化文本转化为定量可靠性指标。

Comments An adjustable template containing the Python script architecture, applied dynamic prompts, and data schemas is hosted in an open-source GitHub repository: https://github.com/mvmalyi/llm-driven-wind-turbine-maintenance-log-labelling

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27152 2026-08-14 cs.SI 版本更新 90%

Disrupting Networks: Amplifying Social Dissensus via Opinion Perturbation and Large Language Models

干扰网络:通过观点扰动和大语言模型放大社会分歧

Erica Coppolillo, Giuseppe Manco

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 该研究基于Friedkin-Johnsen模型,利用强化学习框架微调大语言模型生成干扰性文本,实现了对社交网络的战略性干扰,其结果对内容审核等领域有重要参考价值。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.10613 2026-08-12 cs.SE 新提交 90%

CausalRepair: Bridging the Causality Gap in Large Language Model-Based Automated Program Repair via Dual-Slicing

CausalRepair:通过双重切片弥合基于大语言模型的自动化程序修复中的因果差距

Linhao Wu, Yizhou Chen, Zhen Yang, Pengyu Xue, Dan Hao

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 CausalRepair是基于最小因果上下文的对话驱动型自动化程序修复框架,通过双重切片策略构建因果相关上下文,在Defects4J数据集上修复313个漏洞,性能优于现有方法且修复成本更低。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.07488 2026-08-11 cs.HC cs.CY 新提交 90%

Large Language Models Explain Experts Better Than Experts Themselves

大型语言模型(LLM)比专家自身更擅长解释专家

Mina Cho, Russell J. Funk, Alok Gupta, Mochen Yang

专题命中 其他LLM :LLM(title_cn,abstract);large language model(title);language model(title)

AI总结 该研究表明,大型语言模型可从专家行为中外部化隐性知识,提升决策质量并帮助新手接近专家表现,为波兰尼悖论提供实证支持,凸显其作为克服专家表述瓶颈的可扩展工具的潜力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.22356 2026-08-07 cs.HC 版本更新 90%

Large Language Model Counterarguments in Older Adults: Cognitive Offloading or Susceptibility to Moral Persuasion?

大语言模型在老年人中的反论点:认知卸载还是对道德说服的脆弱性?

Kou Tamura, Sayaka Ishibashi, Ayana Goma, Kenta Yamamoto, Kouhei Masumoto

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 研究发现大语言模型能显著影响老年人和年轻人的道德判断,老年人更易受说服,认知功能较低者在情感冲突困境中更易接受反论点,表明LLMs可能成为认知卸载工具,但也对认知脆弱者构成风险。

Comments This paper has been published in Computers in Human Behavior. The final published version is available at https://doi.org/10.1016/j.chb.2026.109142

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.02420 2026-08-04 cs.HC cs.CY 新提交 90%

WIP: Chat-Debugging: Large Language Model as a Hardware Debugging Assistant

WIP:Chat-Debugging:大型语言模型作为硬件调试助手

Andrew Ash, John Hu

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 该研究提出将大型语言模型应用于硬件调试的Chat-Debugging方法,通过人机交互提升电气类学生的调试信心与技能,填补了物理硬件调试辅助工具的研究空白。

Comments This is the accepted version of a paper accepted for presentation at the 2026 IEEE Frontiers in Education Conference (FIE). The final version will be available via IEEE Xplore at: https://ieeexplore.ieee.org/Xplore/home.jsp

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.22657 2026-07-30 cs.NE 版本更新 90%

Automatic programming via large language models with population self-evolution for dynamic fuzzy job shop scheduling problem

结合种群自进化大语言模型的动态模糊作业车间调度问题自动编程

Jin Huang, Qihao Liu, Xinyu Li, Liang Gao, Yue Teng

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 该研究针对动态模糊作业车间调度问题,提出种群自进化框架结合大语言模型自动设计启发式调度规则,性能优于多种现有方法。

Comments 13 pages, 10 figures. Accepted for publication in IEEE Transactions on Fuzzy Systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.03202 2026-07-21 cs.CY 版本更新 90%

Prosocial Persuasion at Scale? Large Language Models Outperform Humans in Donation Appeals Across Levels of Personalization

大规模的亲社会说服?大语言模型在不同个性化水平上的捐赠呼吁中优于人类

John Caffier, Olga Stavrova, Bennett Kleinberg

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 研究探讨大语言模型生成的捐赠呼吁在不同个性化水平下的有效性,发现其在捐款数额、参与度和说服力上均优于人类创作内容,但虚假个性化会带来负面影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.12235 2026-07-15 cs.CY 新提交 90%

A Semi-Automated System for Generating Dialogue-Based TTS Lessons Using Large Language Models: An Exploratory Study of Educational Potential

一种使用大语言模型生成基于对话的语音合成课程的半自动系统:教育潜力的探索性研究

Gendo Kumoi, Fumie Watanabe, Tota Suko, Takashi Ishida, Yuko Kuma, Manabu Kobayashi, Shigeichi Hirasawa

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 研究提出用大语言模型和文本转语音技术生成基于对话课程的半自动系统,经准实验探索其教育潜力。系统通过三阶段工作流程增强教育工作者,引入新方法。实验表明对话TTS在理解等方面优于单声道TTS,为TTS音频教育可接受性及课程形式设计提供依据。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.28127 2026-06-29 cs.CL cs.AI cs.LG 新提交 90%

From Tokens to States: LLMs as a Special Case of World Models and the Continuous Path Beyond

从令牌到状态:LLM作为世界模型的特例及其连续路径

Paul Dubois

机构 * Paul Dubois(保罗·杜博伊斯)

专题命中 其他LLM :LLM(title_cn,summary_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文论证LLM是世界模型的退化特例,并提出从NTP到JEPA的连续谱系,逐步放松LLM约束,同时探讨其可扩展性挑战。

Comments 10 pages, 6 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09303 2026-05-28 q-fin.PM 90%

Investor risk profiles of large language models

大型语言模型的投资者风险画像

Hanyong Cho, Geumil Bae, Jang Ho Kim

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 本研究通过标准化风险问卷评估GPT、Gemini和Llama三种大型语言模型的风险偏好,发现它们普遍为长期投资者但风险容忍度不同,且赋予特定人物角色后各模型会调整其风险画像。

Comments Poster presented at the AI for Finance Symposium '25, The 6th ACM International Conference on AI in Finance (ICAIF '25)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15840 2026-04-28 cs.LO cs.FL 90%

Automatic Generation of Safety-compliant Linear Temporal Logic via Large Language Model: A Self-supervised Framework

基于大语言模型的自动安全合规线性时序逻辑生成:一种自监督框架

Junle Li, Siqi Chen, Jiakai Li, Meiqi Tian, Bingzhuo Zhong

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 本文提出AutoSafeLTL框架,利用大语言模型自动生成符合安全限制的LTL规范,通过语言包含检查与自动反例引导修改机制确保逻辑一致性和语义准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.04288 2026-04-16 cs.CR cs.SE 90%

LLM-Enabled Open-Source Systems in the Wild: An Empirical Study of Vulnerabilities in GitHub Security Advisories

基于大语言模型的开源系统在现实中的应用:对GitHub安全通告中漏洞的实证研究

Fariha Tanjim Shifat, Hariswar Baburaj, Ce Zhou, Jaydeb Sarker, Mia Mohammad Imran

专题命中 其他LLM :LLM(title,summary_cn);large language model(abstract,comments);language model(abstract,comments)

AI总结 本研究分析了295份GitHub安全通告,发现大多数漏洞映射到已知CWE,但模型中介暴露不足,建议结合CWE和OWASP视角更全面地评估LLM集成系统的漏洞。

Comments The 2nd International Workshop on Large Language Model Supply Chain Analysis (LLMSC 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.14826 2026-04-07 cs.CL cs.AI cs.HC cs.LG 90%

SPRIG: Improving Large Language Model Performance by System Prompt Optimization

SPRIG:通过系统提示优化提升大语言模型性能

Lechen Zhang, Tolga Ergen, Lajanugen Logeswaran, Moontae Lee, David Jurgens

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) University of Michigan(密歇根大学) University of Illinois Chicago(伊利诺伊大学芝加哥分校) LG AI Research(LG AI 研究院)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出SPRIG,一种基于编辑的遗传算法,通过优化通用提示提升大语言模型性能,发现系统提示与任务提示结合可进一步提升效果。

Comments Accepted at ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.15437 2026-01-23 cs.HC 90%

Exploring Implicit Perspectives on Autism in Large Language Models Through Multi-Agent Simulations

通过多智能体模拟探索大型语言模型中对自闭症的隐含视角

Sohyeon Park, Jesus Armando Beltran, Aehong Min, Anamara Ritt-Olson, Gillian R. Hayes

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

AI总结 通过多智能体模拟研究大型语言模型对自闭症的隐含视角,揭示其偏见并提出双重共情问题以改进与自闭症人群的互动。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.20408 2026-01-08 q-bio.NC cs.AI cs.CL cs.LG 90%

Brain-Inspired Exploration of Functional Networks and Key Neurons in Large Language Models

受脑启发的大型语言模型中功能网络和关键神经元探索

Yiheng Liu, Zhengliang Liu, Zihao Wu, Junhao Ning, Haiyang Sun, Sichen Xia, Yang Yang, Xiaohui Gao, Ning Qiang, Bao Ge, Tianming Liu, Junwei Han, Xintao Hu

机构 * School of Automation, Northwestern Polytechnical University, Xi’an, China(自动化学院,西北工业大学,西安,中国) School of Computing, University of Georgia, Athens, USA(计算机学院,佐治亚大学,亚特兰大,美国) School of Physics and Information Technology, Shaanxi Normal University, Xi’an, China(物理与信息技术学院,陕西师范大学,西安,中国)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文受脑启发,探索LLM中的功能网络和关键神经元,发现这些网络对模型性能至关重要,通过抑制或增强网络活动可影响模型整体表现或特定任务效果。

Comments 21 pages, 18 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.06192 2025-12-03 cs.DB cs.AI cs.CL cs.LG 90%

SQLBarber: A System Leveraging Large Language Models to Generate Customized and Realistic SQL Workloads

SQLBarber: 借助大型语言模型生成定制化和真实SQL工作负载的系统

Jiale Lao, Immanuel Trummer

机构 * Cornell University(康奈尔大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 SQLBarber利用大型语言模型生成定制化且真实的SQL工作负载,通过声明式接口和贝叶斯优化器高效生成符合目标成本分布的查询。

Comments Accepted by SIGMOD 2026; extended version with appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.14033 2025-11-18 cs.AI cs.CL cs.LG 90%

MLR-Copilot: Autonomous Machine Learning Research based on Large Language Models Agents

Ruochen Li, Teerth Patel, Qingyun Wang, Xinya Du

机构 * University of Texas at Dallas(德克萨斯大学达拉斯分校) UIUC(伊利诺伊大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12830 2025-08-15 cs.CL cs.AI cs.LG 90%

Knowledge-based Consistency Testing of Large Language Models

Sai Sathiesh Rajan, Ezekiel Soremekun, Sudipta Chattopadhyay

机构 * Singapore University of Technology and Design, Singapore(新加坡科技设计大学) Royal Holloway, University of London, UK(伦敦大学皇家霍洛威学院)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 12 pages, 4 figures, 8 tables, Accepted at EMNLP 2024 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12856 2025-08-12 cs.CL cs.AI cs.CY cs.LG 90%

AI-AI Bias: large language models favor communications generated by large language models

Walter Laurito, Benjamin Davis, Peli Grietzer, Tomáš Gavenčiak, Ada Böhm, Jan Kulveit

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 8 pages, 4 figures

Journal ref Proc. Natl. Acad. Sci. U.S.A. 122 (31) e2415697122 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.11316 2025-07-16 cs.CL cs.AI cs.LG 90%

Internal Value Alignment in Large Language Models through Controlled Value Vector Activation

Haoran Jin, Meng Li, Xiting Wang, Zhihao Xu, Minlie Huang, Yantao Jia, Defu Lian

机构 * University of Science and Technology of China(中国科学技术大学) State Key Laboratory of Cognitive Intelligence(认知智能国家重点实验室) Gaoling School of Artificial Intelligence(光明人工智能学院) Renmin University of China(中国人民大学) Beijing Key Laboratory of Research on Large Models and Intelligent Governance(北京大型模型与智能治理重点实验室) Engineering Research Center of Next-Generation Intelligent Search and Recommendation, MOE(下一代智能搜索与推荐工程技术研究中心) Tsinghua University(清华大学) Huawei Technologies Co. Ltd(华为技术有限公司)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 25 pages, 14 figures. Accepted by ACL 2025 (main conference)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.03247 2025-06-13 cs.HC 90%

End User Authoring of Personalized Content Classifiers: Comparing Example Labeling, Rule Writing, and LLM Prompting

Leijie Wang, Kathryn Yurechko, Pranati Dani, Quan Ze Chen, Amy X. Zhang

专题命中 其他LLM :LLM(title,abstract);prompting(title,abstract);large language model(abstract);language model(abstract)

Comments Accepted by CHI'25

详情

展开后加载摘要…

URL PDF HTML 收藏