arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-10-28 至 2025-10-28 共收录 412 信号源:cs.CL, cs.AI, cs.LG

1. 效率与部署 82 篇

2510.22317 2025-10-28 cs.CL 86%

Memory-based Language Models: An Efficient, Explainable, and Eco-friendly Approach to Large Language Modeling

Antal van den Bosch, Ainhoa Risco Patón, Teun Buijse, Peter Berck, Maarten van Gompel

机构 * Utrecht University(乌特勒支大学) Lund University(吕勒奥大学) Royal Netherlands Academy of Arts and Sciences(荷兰艺术与科学皇家学院)

专题命中 效率与部署 :language model(title,abstract);large language model(title);分类 cs.CL

Comments 15 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.23408 2025-10-28 cs.AI cs.DC cs.ET cs.LG cs.MA 86%

AutoStreamPipe: LLM Assisted Automatic Generation of Data Stream Processing Pipelines

Abolfazl Younesi, Zahra Najafabadi Samani, Thomas Fahringer

机构 * Departement of Computer Science, University of Innsbruck(因斯布鲁克大学计算机科学系)

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24749 2025-10-28 cs.LG cs.CL math.OC 86%

SUMO: Subspace-Aware Moment-Orthogonalization for Accelerating Memory-Efficient LLM Training

Yehonathan Refael, Guy Smorodinsky, Tom Tirer, Ofir Lindenbaum

机构 * Faculty of Engineering(工程学院) Tel Aviv University(特拉维夫大学) Department of Computer science(计算机科学系) Ben Gurion University(本· Gurion大学) Faculty of Engineering Bar-Ilan University(巴伊兰大学工程学院) Bar-Ilan University(巴伊兰大学)

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Journal ref The Thirty-Ninth Annual Conference on Neural Information Processing Systems (NeurIPS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21883 2025-10-28 cs.CL cs.AI 86%

Language Ranker: A Lightweight Ranking framework for LLM Decoding

Chenheng Zhang, Tianqi Du, Jizhe Zhang, Mingqing Xiao, Yifei Wang, Yisen Wang, Zhouchen Lin

机构 * State Key Lab of General Artificial Intelligence, School of Intelligence Science and Technology, Peking University(一般人工智能国家重点实验室,智能科学与技术学院,北京大学) Institute for Artificial Intelligence, Peking University(人工智能研究院,北京大学) MIT CSAIL, MA, USA(MIT CSAIL,马萨诸塞州,美国) Microsoft Research Asia(微软亚洲研究院)

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.13000 2025-10-28 cs.CL cs.AI cs.IR 86%

RAGE Against the Machine: Retrieval-Augmented LLM Explanations

Joel Rorseth, Parke Godfrey, Lukasz Golab, Divesh Srivastava, Jaroslaw Szlichta

机构 * University of Waterloo(滑铁卢大学) York University(约克大学)

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accepted by ICDE 2024 (Demonstration Track)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.21433 2025-10-28 cs.SE cs.AI 85%

The Complexity Trap: Simple Observation Masking Is as Efficient as LLM Summarization for Agent Context Management

Tobias Lindenbauer, Igor Slinko, Ludwig Felder, Egor Bogomolov, Yaroslav Zharov

机构 * JetBrains Research(JetBrains研究院) School of Computation, Information and Technology, Technical University of Munich(技术大学慕尼黑计算、信息与技术学院)

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments v3: DL4C camera-ready version to be presented at the 4th DL4C workshop co-located with NeurIPS '25; added OpenHands generality probe, added hybrid context management strategy

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22784 2025-10-28 cs.RO cs.AI 85%

PIP-LLM: Integrating PDDL-Integer Programming with LLMs for Coordinating Multi-Robot Teams Using Natural Language

Guangyao Shi, Yuwei Wu, Vijay Kumar, Gaurav S. Sukhatme

机构 * Department of Computer Science, University of Southern California(计算机科学系,南加州大学) GRASP Lab, University of Pennsylvania(GRASP实验室,宾夕法尼亚大学)

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22256 2025-10-28 cs.CL 85%

SteerX: Disentangled Steering for LLM Personalization

Xiaoyan Zhao, Ming Yan, Yilun Qiu, Haoting Ni, Yang Zhang, Fuli Feng, Hong Cheng, Tat-Seng Chua

机构 * The Chinese University of Hong Kong, University of Science and Technology of China, National University of Singapore(香港中文大学、中国科学技术大学、新加坡国立大学)

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10644 2025-10-28 cs.AI 85%

Hierarchical Optimization via LLM-Guided Objective Evolution for Mobility-on-Demand Systems

Yi Zhang, Yushen Long, Yun Ni, Liping Huang, Xiaohong Wang, Jun Liu

机构 * Agency for Science, Technology and Research(科技研究局) Morgan Stanley Asia Pte.(摩根大通亚洲公司) Onto Innovation Inc.(Onto Innovation公司) School of computing and communications, Lancaster University(兰卡斯特大学计算机与通讯学院)

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.02890 2025-10-28 cs.CR cs.IT cs.LG math.IT 85%

Theoretically Grounded Framework for LLM Watermarking: A Distribution-Adaptive Approach

Haiyun He, Yepeng Liu, Ziqiao Wang, Yongyi Mao, Yuheng Bu

机构 * HKUST (GZ)(香港科技大学(广州)) UC Santa Barbara(加州大学圣芭芭拉分校) Tongji University(同济大学) University of Ottawa(渥太华大学)

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.09501 2025-10-28 cs.CL 85%

Understanding and Mitigating Numerical Sources of Nondeterminism in LLM Inference

Jiayi Yuan, Hao Li, Xinheng Ding, Wenya Xie, Yu-Jhe Li, Wentian Zhao, Kun Wan, Jing Shi, Xia Hu, Zirui Liu

机构 * Rice University(Rice大学) University of Minnesota Twin Cities(明尼苏达大学双城分校) Adobe Inc.(Adobe公司)

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.20524 2025-10-28 cs.LG 85%

Towards Fully FP8 GEMM LLM Training at Scale

Alejandro Hernández-Cano, Dhia Garbaya, Imanol Schlag, Martin Jaggi

机构 * EPFL(苏黎世联邦理工学院) ETHZ(苏黎世联邦理工学院)

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.LG

Comments 19 pages including appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.00775 2025-10-28 cs.HC 85%

Efficiency with Rigor! A Trustworthy LLM-powered Workflow for Qualitative Data Analysis

Jie Gao, Zhiyao Shu, Shun Yi Yeo, Alok Prakash, Chien-Ming Huang, Mark Dredze, Ziang Xiao

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21794 2025-10-28 cs.CV cs.AI 83%

Token-Level Inference-Time Alignment for Vision-Language Models

Kejia Chen, Jiawen Zhang, Jiacong Hu, Kewei Gao, Jian Lou, Zunlei Feng, Mingli Song

机构 * Zhejiang University(浙江大学) Sun Yat-sen University(中山大学)

专题命中 效率与部署 :language model(title,abstract);preference optimization(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22852 2025-10-28 cs.LG cs.AI 81%

Encoder-Decoder Diffusion Language Models for Efficient Training and Inference

Marianne Arriola, Yair Schiff, Hao Phung, Aaron Gokaslan, Volodymyr Kuleshov

机构 * Department of Computer Science, Cornell University(计算机科学系,康奈尔大学)

专题命中 效率与部署 :language model(title,abstract);分类 cs.AI、cs.LG

Comments NeurIPS 2025. We provide the code at https://github.com/kuleshov-group/e2d2

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22641 2025-10-28 cs.LG cs.AI 81%

FastVLM: Self-Speculative Decoding for Fast Vision-Language Model Inference

Divya Jyoti Bajpai, Manjesh Kumar Hanawal

机构 * Dept. of IEOR, IIT Bombay(工业工程与运营研究系,印度理工学院班加罗尔)

专题命中 效率与部署 :language model(title,abstract);分类 cs.AI、cs.LG

Comments Accepted for presentation at the main Conference IJCNLP-AACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22451 2025-10-28 cs.LG cs.AI 81%

GraphTOP: Graph Topology-Oriented Prompting for Graph Neural Networks

Xingbo Fu, Zhenyu Lei, Zihan Chen, Binchi Zhang, Chuxu Zhang, Jundong Li

机构 * University of Virginia(弗吉尼亚大学) University of Connecticut Storrs(康涅狄格大学斯特劳斯分校)

专题命中 效率与部署 :prompting(title,abstract);分类 cs.AI、cs.LG

Comments Accepted by the 39 Annual Conference on Neural Information Processing Systems (NeurIPS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22257 2025-10-28 cs.LG cs.AI 81%

LUNA: Efficient and Topology-Agnostic Foundation Model for EEG Signal Analysis

Berkay Döner, Thorir Mar Ingolfsson, Luca Benini, Yawei Li

机构 * Integrated Systems Laboratory, ETH Zürich(瑞士苏黎世联邦理工学院集成系统实验室)

专题命中 效率与部署 :foundation model(title,abstract);分类 cs.AI、cs.LG

Comments NeurIPS camera-ready version, 27 pages, 10 figures, 13 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.17660 2025-10-28 cs.LG 79%

Effortless, Simulation-Efficient Bayesian Inference using Tabular Foundation Models

Julius Vetter, Manuel Gloeckler, Daniel Gedon, Jakob H. Macke

机构 * Machine Learning in Science, University of Tübingen(图宾根大学机器学习与科学系) Tübingen AI Center(图宾根人工智能中心) Department Empirical Inference, Max Planck Institute for Intelligent Systems(智能系统研究所经验推断部门)

专题命中 效率与部署 :foundation model(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.08441 2025-10-28 cs.CV cs.AI 79%

Vision Foundation Models as Effective Visual Tokenizers for Autoregressive Image Generation

Anlin Zheng, Xin Wen, Xuanyang Zhang, Chuofan Ma, Tiancai Wang, Gang Yu, Xiangyu Zhang, Xiaojuan Qi

机构 * The University of Hong Kong(香港大学) StepFun Dexmal MEGVII Technology(MEGVII科技)

专题命中 效率与部署 :foundation model(title,abstract);分类 cs.AI

Comments 20 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.09663 2025-10-28 cs.LG 79%

Analog Foundation Models

Julian Büchel, Iason Chalas, Giovanni Acampa, An Chen, Omobayode Fagbohungbe, Sidney Tsai, Kaoutar El Maghraoui, Manuel Le Gallo, Abbas Rahimi, Abu Sebastian

机构 * IBM Research – Zurich(IBM瑞士研究实验室) ETH Zürich(苏黎世联邦理工学院) IBM Research – Almaden(IBM加州阿尔玛登研究实验室) IBM Thomas J. Watson Research Center(IBM托马斯·J·沃森研究中心)

专题命中 效率与部署 :foundation model(title,abstract);分类 cs.LG

Comments Neural Information Processing Systems (NeurIPS) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.04262 2025-10-28 cs.LG stat.ME stat.ML 79%

Efficient Randomized Experiments Using Foundation Models

Piersilvio De Bartolomeis, Javier Abad, Guanbo Wang, Konstantin Donhauser, Raymond M. Duch, Fanny Yang, Issa J. Dahabreh

专题命中 效率与部署 :foundation model(title,abstract);分类 cs.LG

Comments Accepted for presentation at the Conference on Neural Information Processing Systems (NeurIPS) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21879 2025-10-28 cs.CV cs.AI 79%

TernaryCLIP: Efficiently Compressing Vision-Language Models with Ternary Weights and Distilled Knowledge

Shu-Hao Zhang, Wei-Cheng Tang, Chen Wu, Peng Hu, Nan Li, Liang-Jie Zhang, Qi Zhang, Shao-Qun Zhang

机构 * State Key Laboratory of Novel Software Technology, Nanjing University(新型软件技术国家重点实验室) Microsoft AI(微软人工智能)

专题命中 效率与部署 :language model(title);pretraining(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22679 2025-10-28 cs.AI cs.CL 79%

Do Stop Me Now: Detecting Boilerplate Responses with a Single Iteration

Yuval Kainan, Shaked Zychlinski

机构 * JFrog

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 13 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22475 2025-10-28 cs.CL cs.AI 79%

CHOIR: Collaborative Harmonization fOr Inference Robustness

Xiangjue Dong, Cong Wang, Maria Teleki, Millennium Bismay, James Caverlee

机构 * Texas A&M University(德克萨斯大学)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments updated version

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.23341 2025-10-28 cs.CL 77%

LightKGG: Simple and Efficient Knowledge Graph Generation from Textual Data

Teng Lin

机构 * DSA,HKUST(GZ)(香港科技大学(广州))

专题命中 效率与部署 :large language model(abstract);language model(abstract);SLM(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.05530 2025-10-28 cs.DB cs.LG cs.PF 77%

Leveraging Approximate Caching for Faster Retrieval-Augmented Generation

Shai Bergman, Anne-Marie Kermarrec, Diana Petrescu, Rafael Pires, Mathis Randl, Martijn de Vos, Ji Zhang

机构 * Huawei Research(华为研究)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

Comments Accepted at Middleware '25

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.13317 2025-10-28 cs.LG 77%

Unlabeled Data vs. Pre-trained Knowledge: Rethinking SSL in the Era of Large Models

Song-Lin Lv, Rui Zhu, Tong Wei, Yu-Feng Li, Lan-Zhe Guo

机构 * School of Intelligence Science and Technology, Nanjing University, China(智能科学与技术学院,南京大学) School of Artificial Intelligence, Nanjing University, China(人工智能学院,南京大学) National Key Laboratory for Novel Software Technology, Nanjing University, China(新型软件技术国家重点实验室,南京大学) School of Computer Science and Engineering, Southeast University, Nanjing, China(计算机科学与工程学院,东南大学)

专题命中 效率与部署 :large language model(abstract);language model(abstract);foundation model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.03112 2025-10-28 cs.LG 77%

Plug-and-Play AMC: Context Is King in Training-Free, Open-Set Modulation with LLMs

Mohammad Rostami, Atik Faysal, Reihaneh Gh. Roshan, Huaxia Wang, Nikhil Muralidhar, Yu-Dong Yao

机构 * Electrical and Computer Engineering(电气与计算机工程) Rowan University(罗文大学) Computer Science(计算机科学) Stevens Institute of Technology(史蒂文斯理工学院)

专题命中 效率与部署 :LLM(abstract);language model(abstract);foundation model(abstract);分类 cs.LG

Journal ref 2025 IEEE 34th Wireless and Optical Communications Conference (WOCC), pp. 345-350

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.14305 2025-10-28 cs.IR cs.LG 77%

Scaling Down, Serving Fast: Compressing and Deploying Efficient LLMs for Recommendation Systems

Kayhan Behdin, Ata Fatahibaarzi, Qingquan Song, Yun Dai, Aman Gupta, Zhipeng Wang, Shao Tang, Hejian Sang, Gregory Dexter, Sirou Zhu, Siyu Zhu, Tejas Dharamsi, Vignesh Kothapalli, Zhoutong Fu, Yihan Cao, Pin-Lun Hsu, Fedor Borisyuk, Natesh Pillai, Luke Simon, Rahul Mazumder

机构 * LinkedIn LinkedIn LLM Efficiency Core Team(LinkedIn LLM效率核心团队) MIT(麻省理工学院)

专题命中 效率与部署 :large language model(abstract);language model(abstract);small language model(abstract);分类 cs.LG

Comments Accepted to EMNLP 2025 Industry Track - Oral Presentation

详情

展开后加载摘要…

URL PDF HTML 收藏