arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-10-16 至 2025-10-16 共收录 175 信号源:cs.CL, cs.AI, cs.LG

1. 效率与部署 21 篇

2510.12041 2025-10-16 cs.CL 81%

Improving Text-to-Image Generation with Input-Side Inference-Time Scaling

Ruibo Chen, Jiacheng Pan, Heng Huang, Zhenheng Yang

机构 * TikTok University of Maryland, College Park(马里兰大学)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);preference optimization(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13713 2025-10-16 cs.LG math.OC 77%

Don't Be Greedy, Just Relax! Pruning LLMs via Frank-Wolfe

Christophe Roux, Max Zimmer, Alexandre d'Aspremont, Sebastian Pokutta

机构 * Department for AI in Society, Science, and Technology, Zuse Institute Berlin(人工智能、科学与技术部门,柏林Zuse研究所) Institute of Mathematics, Technische Universität Berlin(数学研究所,柏林技术大学) CNRS & D.I. École Normale Supérieure, Paris, France(法国巴黎国家科学研究中心及École Normale Supérieure)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13401 2025-10-16 cs.AR cs.DC cs.LG 77%

F-BFQ: Flexible Block Floating-Point Quantization Accelerator for LLMs

Jude Haris, José Cano

机构 * School of Computing Science, University of Glasgow(计算科学学院,格拉斯哥大学)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

Comments Accepted to Workshop on New Approaches for Addressing the Computing Requirements of LLMs and GNNs (LG-ARC) @ ISCA 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13011 2025-10-16 cs.HC cs.AI 77%

Deliberate Lab: A Platform for Real-Time Human-AI Social Experiments

Crystal Qian, Vivian Tsai, Michael Behr, Nada Hussein, Léo Laugier, Nithum Thain, Lucas Dixon

机构 * Google DeepMind(谷歌DeepMind) Google(谷歌) EPFL(苏黎世联邦理工学院)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13558 2025-10-16 cs.SD 75%

Steer-MoE: Efficient Audio-Language Alignment with a Mixture-of-Experts Steering Module

Ruitao Feng, Bixi Zhang, Sheng Liang, Zheng Yuan

机构 * The University of Hong Kong, Fauclty of Science, Hong Kong(香港大学科学学院) Aix-Marseille University, Laboratoire Parole et Langage (LPL), France(艾克斯-马赛大学语言与言语实验室(LPL))

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract)

Comments 5 pages, 1 figures. Code is available at: https://github.com/forfrt/SteerMoE. Submitted to ICASSP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.13780 2025-10-16 cs.IT cs.AI cs.CL cs.LG math.IT 75%

Optimal Quantization for Matrix Multiplication

Or Ordentlich, Yury Polyanskiy

机构 * Hebrew University of Jerusalem(耶路撒冷希伯来大学) MIT(麻省理工学院)

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13786 2025-10-16 cs.LG cs.AI 73%

The Art of Scaling Reinforcement Learning Compute for LLMs

Devvrit Khatri, Lovish Madaan, Rishabh Tiwari, Rachit Bansal, Sai Surya Duvvuri, Manzil Zaheer, Inderjit S. Dhillon, David Brandfonbrener, Rishabh Agarwal

机构 * Meta Harvard University(哈佛大学) Periodic Labs

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments 28 pages, 20 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12803 2025-10-16 cs.SE cs.AI cs.CL cs.PL 73%

AutoCode: LLMs as Problem Setters for Competitive Programming

Shang Zhou, Zihan Zheng, Kaiyuan Liu, Zeyu Shen, Zerui Cheng, Zexing Chen, Hansen He, Jianzhu Yao, Huanzhi Mao, Qiuyang Mang, Tianfu Fu, Beichen Li, Dongruixuan Li, Wenhao Chai, Zhuang Liu, Aleksandra Korolova, Peter Henderson, Natasha Jaques, Pramod Viswanath, Saining Xie, Jingbo Shang

机构 * University of California San Diego(加州大学圣地亚哥分校) New York University(纽约大学) University of Washington(华盛顿大学) Princeton University(普林斯顿大学) University of California Berkeley(加州大学伯克利分校) OpenAI Massachusetts Institute of Technology(麻省理工学院) University of Waterloo(滑铁卢大学) Sentient Labs

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Project page: https://livecodebenchpro.com/projects/autocode/overview

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13147 2025-10-16 cs.AR cs.LG cs.PF 70%

D-com: Accelerating Iterative Processing to Enable Low-rank Decomposition of Activations

Faraz Tahmasebi, Michael Pelluer, Hyoukjun Kwon

机构 * University of California, Irvine(加州大学尔湾分校) NVIDIA(NVIDIA公司)

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.LG

Comments 12 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12966 2025-10-16 cs.CL 70%

3-Model Speculative Decoding

Sanghyun Byun, Mohanad Odema, Jung Ick Guack, Baisub Lee, Jacob Song, Woo Seong Chung

机构 * LG Electronics USA(LG电子美国公司)

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted at NeurIPS SPIGM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12825 2025-10-16 cs.CL cs.AI cs.DB cs.LG 67%

Classifier-Augmented Generation for Structured Workflow Prediction

Thomas Gschwind, Shramona Chakraborty, Nitin Gupta, Sameep Mehta

机构 * IBM Research(IBM研究院)

专题命中 效率与部署 :prompting(abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20720 2025-10-16 quant-ph 67%

Accelerating the drive towards energy-efficient generative AI with quantum computing algorithms

Frederik F. Flöther, Jan Mikolon, Maria Longobardi

专题命中 效率与部署 :large language model(abstract);language model(abstract)

Journal ref Frederik F Flöther et al 2025 Quantum Sci. Technol. 10 040501

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.19847 2025-10-16 cs.LG cs.AI cs.CL cs.CV 67%

Orthogonal Finetuning Made Scalable

Zeju Qiu, Weiyang Liu, Adrian Weller, Bernhard Schölkopf

机构 * Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所) The Chinese University of Hong Kong(香港中文大学) University of Cambridge(剑桥大学) The Alan Turing Institute(艾伦·图灵研究所)

专题命中 效率与部署 :foundation model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments EMNLP 2025 Main (18 pages, 7 figures, project page: https://spherelab.ai/oftv2/)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14755 2025-10-16 physics.ed-ph cs.AI 57%

Reliable generation of isomorphic physics problems using Generative AI with prompt-chaining and tool use

Zhongzhou Chen

专题命中 效率与部署 :LLM(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.16992 2025-10-16 cs.CV cs.LG 57%

Segment Anything Model is a Good Teacher for Local Feature Learning

Jingqian Wu, Rongtao Xu, Zach Wood-Doughty, Changwei Wang, Shibiao Xu, Edmund Y. Lam

机构 * The University of Hong Kong(香港大学) State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,中国科学院自动化研究所) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) Northwestern University(西北大学) Key Laboratory of Computing Power Network and Information Security, Ministry of Education, Shandong Computer Science Center (National Supercomputer Center in Jinan), Qilu University of Technology (Shandong Academy of Sciences)(计算能力网络与信息安全重点实验室,教育部长江计算机科学中心(济南国家超算中心),齐鲁工业大学(山东科学院)) Shandong Provincial Key Laboratory of Computer Networks, Shandong Fundamental Research Center for Computer Science(山东省计算机网络重点实验室,山东省计算机科学基础研究中心)

专题命中 效率与部署 :prompting(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13678 2025-10-16 cs.CV 50%

FlashWorld: High-quality 3D Scene Generation within Seconds

Xinyang Li, Tengfei Wang, Zixiao Gu, Shengchuan Zhang, Chunchao Guo, Liujuan Cao

机构 * Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University(中国教育部多媒体可信感知与高效计算重点实验室,厦门大学) Tencent(腾讯) Yes Lab, Fudan University(复旦大学Yes实验室)

专题命中 效率与部署 :post-training(abstract)

Comments Project Page: https://imlixinyang.github.io/FlashWorld-Project-Page/

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 领域大模型 17 篇

2510.13255 2025-10-16 cs.CL cs.NE 89%

Hierarchical Frequency Tagging Probe (HFTP): A Unified Approach to Investigate Syntactic Structure Representations in Large Language Models and the Human Brain

Jingmin An, Yilong Song, Ruolin Yang, Nai Ding, Lingxi Lu, Yuxuan Wang, Wei Wang, Chu Zhuang, Qian Wang, Fang Fang

机构 * Peking University(北京大学) Zhejiang University(浙江大学) Beijing Language and Culture University(北京语言大学) Beijing Institute for General Artificial Intelligence(北京通用人工智能研究院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.19649 2025-10-16 cs.LG cs.AR 88%

Intelligent4DSE: Optimizing High-Level Synthesis Design Space Exploration with Graph Neural Networks and Large Language Models

Lei Xu, Shanshan Wang, Emmanuel Casseau, Chenglong Xiao

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13371 2025-10-16 cs.IR cs.AI 85%

MADREC: A Multi-Aspect Driven LLM Agent for Explainable and Adaptive Recommendation

Jiin Park, Misuk Kim

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments 18 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13229 2025-10-16 cs.IR 85%

Beyond Static LLM Policies: Imitation-Enhanced Reinforcement Learning for Recommendation

Yi Zhang, Lili Xie, Ruihong Qiu, Jiajun Liu, Sen Wang

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract)

Comments ICDM 2025 Accepted Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12816 2025-10-16 cs.IR 85%

Maximum In-Support Return Modeling for Dynamic Recommendation with Language Model Prior

Xiaocong Chen, Siyu Wang, Lina Yao

专题命中 领域大模型 :language model(title,abstract);LLM(abstract);large language model(abstract)

Comments CIKM'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.10803 2025-10-16 cs.AI cs.CL cs.ET cs.IR 82%

Automated Thematic Analyses Using LLMs: Xylazine Wound Management Social Media Chatter Use Case

JaMor Hairston, Ritvik Ranjan, Sahithi Lakamana, Anthony Spadaro, Selen Bozkurt, Jeanmarie Perrone, Abeed Sarker

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments Pages: 19, Abstract word count: 151 words, Manuscript word count: 2185 words, References: 14, Figures: 3, Tables: 2

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.02379 2025-10-16 cs.CV 82%

MedDINOv3: How to adapt vision foundation models for medical image segmentation?

Yuheng Li, Yizhou Wu, Yuxiang Lai, Mingzhe Hu, Xiaofeng Yang

机构 * Georgia Institute of Technology(佐治亚理工学院) Emory University(埃默里大学) Emory University School of Medicine(埃默里大学医学院)

专题命中 领域大模型 :foundation model(title,abstract);pretraining(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13621 2025-10-16 cs.CY cs.AI 79%

The Role of Computing Resources in Publishing Foundation Model Research

Yuexing Hao, Yue Huang, Haoran Zhang, Chenyang Zhao, Zhenwen Liang, Paul Pu Liang, Yue Zhao, Lichao Sun, Saleh Kalantari, Xiangliang Zhang, Marzyeh Ghassemi

机构 * EECS, MIT(MIT电子工程与计算机科学系) Cornell University(康奈尔大学) CSE, University of Notre Dame(诺丁汉大学计算机科学工程系) Computer Science Department, University of California, Los Angeles(加州大学洛杉矶分校计算机科学系) Computer Science Department, Lehigh University(莱斯大学计算机科学系) School of Advanced Computing, University of Southern California(南加州大学高级计算学院)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13359 2025-10-16 cs.IR cs.CV cs.LG 79%

Improving Visual Recommendation on E-commerce Platforms Using Vision-Language Models

Yuki Yada, Sho Akiyama, Ryo Watanabe, Yuta Ueno, Yusuke Shido, Andre Rusli

机构 * Mercari, Inc.(Mercari公司)

专题命中 领域大模型 :language model(title,abstract);分类 cs.LG

Comments Accepted to ACM RecSys 2025 (Spotlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24101 2025-10-16 cs.CL 77%

BTC-SAM: Leveraging LLMs for Generation of Bias Test Cases for Sentiment Analysis Models

Zsolt T. Kardkovacs, Lynda Djennane, Anna Field, Boualem Benatallah, Yacine Gaci, Fabio Casati, Walid Gaaloul

机构 * Insight SFI Research Center on Data Analytics, Dublin City University(数据分析洞察SFI研究中心,都柏林城市大学) Laboratoire LITAN, École supérieure en Sciences et Technologies de l’Informatique et du Numérique(LITAN实验室,信息与数字技术高等学院) School of Computing, Dublin City University(计算学院,都柏林城市大学) Plus Que Pro, Strasbourg, France(斯特拉斯堡法国Plus Que Pro公司) ServiceNow, Zurich, Switzerland(瑞士苏黎世ServiceNow公司) Department of Information Engineering and Computer Science, University of Trento(信息工程与计算机科学系,特伦托大学) Télécom SudParis, SAMOVAR, Institut Polytechnique de Paris(巴黎理工学院SAMOVAR,Telecom SudParis)

专题命中 领域大模型 :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.CL

Comments Accepted at EMNLP 2025 main conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13575 2025-10-16 cs.SE 75%

Auto-repair without test cases: How LLMs fix compilation errors in large industrial embedded code

Han Fu, Sigrid Eldh, Kristian Wiklund, Andreas Ermedahl, Philipp Haller, Cyrille Artho

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract)

Comments 9 pages, 4 figures, conference: 2025 28th Euromicro Conference on Digital System Design (DSD)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12958 2025-10-16 astro-ph.IM astro-ph.HE astro-ph.SR cs.LG 74%

Simulation-Based Pretraining and Domain Adaptation for Astronomical Time Series with Minimal Labeled Data

Rithwik Gupta, Daniel Muthukrishna, Jeroen Audenaert

机构 * Massachusetts Institute of Technology, Cambridge, MA 02139, USA Irvington High School, Fremont, CA 94538, USA Center for Astrophysics, Harvard \& Smithsonian, Cambridge, MA 02138, USA

专题命中 领域大模型 :pretraining(title);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13115 2025-10-16 cs.CL cs.AI 73%

Multi-Label Clinical Text Eligibility Classification and Summarization System

Surya Tejaswi Yerramsetty, Almas Fathimah

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12920 2025-10-16 astro-ph.IM cs.AI 70%

InferA: A Smart Assistant for Cosmological Ensemble Data

Justin Z. Tam, Pascal Grosset, Divya Banesh, Nesar Ramachandra, Terece L. Turton, James Ahrens

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏