arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-10-03 至 2025-10-03 共收录 38 信号源:cs.CL, cs.AI, cs.LG

1. 效率与部署 38 篇

2510.01863 2025-10-03 cs.NE cs.LG 89%

Microscaling Floating Point Formats for Large Language Models

Marco Cococcioni, Dario Pagani, Federico Rossi

机构 * University of Pisa, Department of Information Engineering(比萨大学信息工程学院)

专题命中 效率与部署 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01240 2025-10-03 cs.LG cs.CL 88%

RSAVQ: Riemannian Sensitivity-Aware Vector Quantization for Large Language Models

Zukang Xu, Xing Hu, Qiang Wu, Dawei Yang

机构 * Houmo AI

专题命中 效率与部署 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17599 2025-10-03 cs.LG 88%

Dynamic Bundling with Large Language Models for Zero-Shot Inference on Text-Attributed Graphs

Yusheng Zhao, Qixin Zhang, Xiao Luo, Weizhi Zhang, Zhiping Xiao, Wei Ju, Philip S. Yu, Ming Zhang

机构 * Peking University(北京大学) Nanyang Technological University(南洋理工大学) University of Washington(华盛顿大学) University of California, Los Angeles(加州大学洛杉矶分校) University of Illinois Chicago(伊利诺伊大学芝加哥分校)

专题命中 效率与部署 :large language model(title,abstract);language model(title,abstract);分类 cs.LG

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18379 2025-10-03 cs.IR 87%

REALM: Recursive Relevance Modeling for LLM-based Document Re-Ranking

Pinhuan Wang, Zhiqiu Xia, Chunhua Liao, Feiyi Wang, Hang Liu

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments EMNLP 2025 (Main Conference, Oral). 15 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01251 2025-10-03 cs.CL stat.ML 85%

Efficient Uncertainty Estimation for LLM-based Entity Linking in Tabular Data

Carlo Bono, Federico Belotti, Matteo Palmonari

机构 * Politecnico di Milano, DEIB(米兰理工学院信息工程学院) Università degli Studi di Milano Bicocca, DISCo(米兰大学比可卡分校信息科学学院)

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01239 2025-10-03 cs.CL 85%

CIFLEX: Contextual Instruction Flow for Sub-task Execution in Multi-Turn Interactions with a Single On-Device LLM

Juntae Lee, Jihwan Bang, Seunghan Yang, Simyung Chang

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments accepted at EMNLP 2025 (main)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01234 2025-10-03 cs.CL 85%

LLMRank: Understanding LLM Strengths for Model Routing

Shubham Agrawal, Prasang Gupta

机构 * Zeno AI

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments 13 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01370 2025-10-03 cs.CV cs.AI cs.LG physics.comp-ph 84%

SPUS: A Lightweight and Parameter-Efficient Foundation Model for PDEs

Abu Bucker Siddik, Diane Oyen, Alexander Most, Michal Kucer, Ayan Biswas

机构 * Computing and Artificial Intelligence Division (CAI)(计算与人工智能部门) Space Remote Sensing and Data Science(空间遥感与数据科学) Los Alamos National Laboratory(洛斯阿拉莫斯国家实验室)

专题命中 效率与部署 :foundation model(title,abstract);pretraining(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01394 2025-10-03 cs.LG cs.CL 82%

Optimal Stopping vs Best-of-$N$ for Inference Time Optimization

Yusuf Kalayci, Vinod Raman, Shaddin Dughmi

机构 * University of Southern California(南加州大学) University of Michigan(密歇根大学)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);RLHF(abstract)

Comments 24 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01645 2025-10-03 cs.CR cs.AI cs.CL cs.LG 80%

Position: Privacy Is Not Just Memorization!

Niloofar Mireshghallah, Tianshi Li

机构 * Carnegie Mellon University(卡内基梅隆大学) Northeastern University(东北大学)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 27 pages, 6 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.02224 2025-10-03 cs.LG stat.ML 79%

Efficiently Generating Correlated Sample Paths from Multi-step Time Series Foundation Models

Ethan Baron, Boris Oreshkin, Ruijun Ma, Hanyu Zhang, Kari Torkkola, Michael W. Mahoney, Andrew Gordon Wilson, Tatiana Konstantinova

机构 * New York University(纽约大学) Amazon(亚马逊)

专题命中 效率与部署 :foundation model(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.01804 2025-10-03 cs.AI 79%

A Study on the MCP x A2A Framework for Enhancing Interoperability of LLM-based Autonomous Agents

Cheonsu Jeong

专题命中 效率与部署 :LLM(title,abstract);分类 cs.AI

Journal ref Journal of Intelligence and Information Systems, 31(3), 141-170 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01910 2025-10-03 cs.LG cs.AI 79%

Are LLMs Better GNN Helpers? Rethinking Robust Graph Learning under Deficiencies with Iterative Refinement

Zhaoyan Wang, Zheng Gao, Arogya Kharel, In-Young Ko

机构 * School of Computing KAIST Daejeon Republic of Korea(韩国国立庆尚大学计算机科学学院) School of Computer Science(计算机科学学院) School of Computing KAIST(韩国国立庆尚大学计算机科学学院)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments 14 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01272 2025-10-03 cs.AI cs.LG 79%

Modeling Others' Minds as Code

Kunal Jha, Aydan Yuenan Huang, Eric Ye, Natasha Jaques, Max Kleiman-Weiner

机构 * Department of Computer Science, University of Washington(华盛顿大学计算机科学系) Department of Computer Science, Johns Hopkins University(约翰霍普金斯大学计算机科学系)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01270 2025-10-03 cs.CL cs.AI 79%

Think Twice, Generate Once: Safeguarding by Progressive Self-Reflection

Hoang Phan, Victor Li, Qi Lei

机构 * New York University(纽约大学)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accepted to EMNLP 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01257 2025-10-03 cs.CL cs.AI 79%

RJE: A Retrieval-Judgment-Exploration Framework for Efficient Knowledge Graph Question Answering with LLMs

Can Lin, Zhengwang Jiang, Ling Zheng, Qi Zhao, Yuhang Zhang, Qi Song, Wangqiu Zhou

机构 * University of Science and Technology of China(中国科学技术大学) City University of Hong Kong(香港城市大学) Deqing Alpha Innovation Institute(德清Alpha创新研究院) Hefei University of Technology(合肥工业大学)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 18 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.11305 2025-10-03 cs.LG cs.AI 79%

QSpec: Speculative Decoding with Complementary Quantization Schemes

Juntao Zhao, Wenhao Lu, Sheng Wang, Lingpeng Kong, Chuan Wu

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Journal ref Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing (EMNLP 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.14497 2025-10-03 cs.CV cs.CL 77%

Efficient Whole Slide Pathology VQA via Token Compression

Weimin Lyu, Qingqiao Hu, Kehan Qi, Zhan Shi, Wentao Huang, Saumya Gupta, Chao Chen

机构 * Stony Brook University(石溪大学)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01670 2025-10-03 cs.AI cs.CL cs.CR cs.CY cs.LG 75%

Just Do It!? Computer-Use Agents Exhibit Blind Goal-Directedness

Erfan Shayegani, Keegan Hines, Yue Dong, Nael Abu-Ghazaleh, Roman Lutz, Spencer Whitehead, Vidhisha Balachandran, Besmira Nushi, Vibhav Vineet

机构 * Microsoft Research AI Frontiers(微软研究院人工智能前沿) Microsoft AI Red Team(微软AI红色团队) University of California, Riverside(加州大学河滨分校) NVIDIA(英伟达)

专题命中 效率与部署 :LLM(abstract);prompting(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.02128 2025-10-03 cs.CL cs.AI 73%

The Disparate Impacts of Speculative Decoding

Jameson Sandler, Ahmet Üstün, Marco Romanelli, Sara Hooker, Ferdinando Fioretto

机构 * University of Virginia(弗吉尼亚大学) Cohere Hofstra University(霍夫斯特大学)

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00499 2025-10-03 cs.CL cs.AI 73%

MOSS-Speech: Towards True Speech-to-Speech Models Without Text Guidance

Xingjian Zhao, Zhe Xu, Qinyuan Cheng, Zhaoye Fei, Luozhijie Jin, Yang Wang, Hanfu Chen, Yaozhou Jiang, Qinghui Gao, Ke Chen, Ruixiao Li, Mingshu Chen, Ruiming Wang, Wenbo Zhang, Yiyang Zhang, Donghua Yu, Yang Gao, Xiaogui Yang, Yitian Gong, Yuanfan Xu, Yaqian Zhou, Xuanjing Huang, Xipeng Qiu

机构 * SII OpenMOSS Team(SII开放MOSS团队) Shanghai Innovation Institute(上海创新研究院) Fudan University(复旦大学)

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23990 2025-10-03 cs.CL cs.AI 73%

The Hidden Costs of Translation Accuracy: Distillation, Quantization, and Environmental Impact

Dhaathri Vijay, Anandaswarup Vadapalli

机构 * University of California, Santa Cruz(加州大学圣克鲁兹分校) Research Spark Hub Inc(Research Spark Hub公司)

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18216 2025-10-03 cs.AI cs.LG 73%

nDNA -- the Semantic Helix of Artificial Cognition

Amitava Das

机构 * BITS Pilani, Goa, India(印度戈阿学院)

专题命中 效率与部署 :foundation model(abstract);pretraining(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14620 2025-10-03 cs.LG cs.CL 73%

Enhancing Learned Knowledge in LoRA Adapters Through Efficient Contrastive Decoding on Ascend NPUs

Morgan Lindsay Heisler, Linzi Xing, Ge Shi, Hanieh Sadri, Gursimran Singh, Weiwei Zhang, Tao Ye, Ying Xiong, Yong Zhang, Zhenan Fan

机构 * Huawei Technologies Canada(华为技术加拿大公司) Huawei Technologies(华为技术)

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments Accepted at ACM KDD 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.10940 2025-10-03 cs.LG cs.AI 73%

CoLA: Compute-Efficient Pre-Training of LLMs via Low-Rank Activation

Ziyue Liu, Ruijie Zhang, Zhengyang Wang, Mingsong Yan, Zi Yang, Paul Hovland, Bogdan Nicolae, Franck Cappello, Sui Tang, Zheng Zhang

机构 * University of California at Santa Barbara(加州大学圣巴巴拉分校) University at Albany, SUNY(阿尔巴尼大学,SUNY) Argonne National Laboratory(阿贡国家实验室)

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments Camera-ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01598 2025-10-03 cs.LG cond-mat.mtrl-sci physics.data-an 70%

Securing generative artificial intelligence with parallel magnetic tunnel junction true randomness

Youwei Bao, Shuhan Yang, Hyunsoo Yang

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.LG

Comments 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01248 2025-10-03 cs.CL 70%

SSTAG: Structure-Aware Self-Supervised Learning Method for Text-Attributed Graphs

Ruyue Liu, Rong Yin, Xiangzhen Bo, Xiaoshuai Hao, Yong Liu, Jinwen Zhong, Can Ma, Weiping Wang

机构 * Institute of Information Engineering, CAS(信息工程研究所,中国科学院) School of Cyberspace Security, UCAS(网络安全学院,中国科学院大学) Wuhan University of Technology(武汉理工大学) Xiaomi EV(小米汽车) Renmin University of China(中国人民大学)

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25779 2025-10-03 cs.AI 70%

Planner-R1: Reward Shaping Enables Efficient Agentic RL with Smaller LLMs

Siyu Zhu, Yanbin Jiang, Hejian Sang, Shao Tang, Qingquan Song, Biao He, Rohit Jain, Zhipeng Wang, Alborz Geramifard

机构 * LinkedIn Corporation(LinkedIn公司)

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23500 2025-10-03 cs.LG 70%

Beyond Outliers: A Study of Optimizers Under Quantization

Georgios Vlassis, Saleh Ashkboos, Alexandra Volkova, Torsten Hoefler, Dan Alistarh

机构 * ETH Zurich(苏黎世联邦理工学院) ISTA(信息科技学院) Red Hat AI(红帽人工智能)

专题命中 效率与部署 :pretraining(abstract);post-training(abstract);分类 cs.LG

Comments 20 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.02104 2025-10-03 cs.RO 67%

LangGrasp: Leveraging Fine-Tuned LLMs for Language Interactive Robot Grasping with Ambiguous Instructions

Yunhan Lin, Wenqi Wu, Zhijie Zhang, Huasong Min

机构 * School of Computer Science and Technology(计算机科学与技术学院) Hubei Province Key Laboratory of Intelligent Information Processing and Real-time Industrial System(湖北省智能信息处理与实时工业系统重点实验室) Institute of Robotics and Intelligent Systems(机器人与智能系统研究院) Wuhan University of Science and Technology(武汉理工大学)

专题命中 效率与部署 :large language model(abstract);language model(abstract)

Comments 8 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏