arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-09-09 至 2025-09-09 共收录 262 信号源:cs.CL, cs.AI, cs.LG

1. 效率与部署 39 篇

2509.05333 2025-09-09 cs.CV cs.AI 79%

RT-VLM: Re-Thinking Vision Language Model with 4-Clues for Real-World Object Recognition Robustness

Junghyun Park, Tuan Anh Nguyen, Dugki Min

专题命中 效率与部署 :language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.08727 2025-09-09 cs.LG cs.AI cs.CY cs.SE 79%

Breaking the ICE: Exploring promises and challenges of benchmarks for Inference Carbon & Energy estimation for LLMs

Samarth Sikand, Rohit Mehra, Priyavanshi Pathania, Nikhil Bamby, Vibhu Saujanya Sharma, Vikrant Kaulgud, Sanjay Podder, Adam P. Burden

机构 * Accenture Labs, India(Accenture实验室,印度) Accenture, India(Accenture,印度) Accenture, USA(Accenture,美国)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments 5 pages. To be published in the proceedings of 9th International Workshop on Green and Sustainable Software (GREENS '25), April 29, 2025, Ottawa, Canada (Co-located with ICSE 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05576 2025-09-09 cs.CV 78%

Sensitivity-Aware Post-Training Quantization for Deep Neural Networks

Zekang Zheng, Haokun Li, Yaofo Chen, Mingkui Tan, Qing Du

机构 * South China University of Technology(华南理工大学) Pazhou Laboratory(Pazhou实验室) Key Laboratory of Big Data and Intelligent Robot, Ministry of Education(大数据与智能机器人重点实验室,教育部)

专题命中 效率与部署 :post-training(title,abstract)

Comments Accepted by PRCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06338 2025-09-09 cs.CR cs.LG 77%

Embedding Poisoning: Bypassing Safety Alignment via Embedding Semantic Shift

Shuai Yuan, Zhibo Zhang, Yuxi Li, Guangdong Bai, Wang Kailong

机构 * University of Electronic Science and Technology of China(电子科技大学) Huazhong University of Science and Technology(华中科技大学) The University of Queensland(昆士兰大学)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

Comments 16 pages,9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.03165 2025-09-09 cs.CL 77%

Efficient Dynamic Clustering-Based Document Compression for Retrieval-Augmented-Generation

Weitao Li, Kaiming Liu, Xiangyu Zhang, Xuanyu Lei, Weizhi Ma, Yang Liu

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05883 2025-09-09 cs.CR cs.AI 77%

Multimodal Prompt Injection Attacks: Risks and Defenses for Modern LLMs

Andrew Yeo, Daeseon Choi

机构 * Ranchview High School(拉文斯维尔高中) Soongsil University(松山大学)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments 8 pages, 4 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06362 2025-09-09 cs.DC 75%

MaaSO: SLO-aware Orchestration of Heterogeneous Model Instances for MaaS

Mo Xuan, Zhang yue, Wu Weigang

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06069 2025-09-09 econ.GN cs.HC q-fin.EC 75%

From Digital Distrust to Codified Honesty: Experimental Evidence on Generative AI in Credence Goods Markets

Alexander Erlei

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18992 2025-09-09 cs.CL cs.AI cs.LG 75%

Automatic Prompt Optimization with Prompt Distillation

Ernest A. Dyagin, Nikita I. Kulin, Artur R. Khairullin, Viktor N. Zhuravlev, Alena N. Sitkina

机构 * Computer Technologies Laboratory(计算机技术实验室) ITMO University(ITMO大学)

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.08755 2025-09-09 cs.CR cs.MA 75%

PILLAR: an AI-Powered Privacy Threat Modeling Tool

Majid Mollaeefar, Andrea Bissoli, Silvio Ranise

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06053 2025-09-09 cs.LG cs.AI 73%

PolicyEvolve: Evolving Programmatic Policies by LLMs for multi-player games via Population-Based Training

Mingrui Lv, Hangzhi Liu, Zhi Luo, Hongjie Zhang, Jie Ou

机构 * School of Computer Science(计算机科学学院) School of Information and Software Engineering(信息与软件工程学院)

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05449 2025-09-09 cs.LG cs.AI 73%

Neural Breadcrumbs: Membership Inference Attacks on LLMs Through Hidden State and Attention Pattern Analysis

Disha Makhija, Manoj Ghuhan Arivazhagan, Vinayshekhar Bannihatti Kumar, Rashmi Gangadharaiah

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11132 2025-09-09 cs.CL 70%

X-EcoMLA: Upcycling Pre-Trained Attention into MLA for Efficient and Extreme KV Compression

Guihong Li, Mehdi Rezagholizadeh, Mingyu Yang, Vikram Appia, Emad Barsoum

机构 * Advanced Micro Devices, Inc.(先进微器件公司)

专题命中 效率与部署 :language model(abstract);post-training(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17746 2025-09-09 cs.CL 70%

Fast Quiet-STaR: Thinking Without Thought Tokens

Wei Huang, Yizhe Xiong, Xin Ye, Zhijie Deng, Hui Chen, Zijia Lin, Guiguang Ding

机构 * School of Computer Science, Beijing University of Posts and Telecommunications(北京邮电大学计算机学院) Tsinghua University(清华大学) Beijing National Research Center for Information Science and Technology (BNRist)(北京信息科学与技术国家研究中心) Kuaishou Technology(快手科技) Shanghai Jiao Tong University(上海交通大学)

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.CL

Comments 10 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03494 2025-09-09 cs.CV 67%

Parameter-Efficient Adaptation of mPLUG-Owl2 via Pixel-Level Visual Prompts for NR-IQA

Yahya Benmahane, Mohammed El Hassouni

机构 * Computer Science Department Faculty of Sciences, Rabat(科学学院计算机科学系,拉巴特) Computer Science Department FLSH(计算机科学系FLSH)

专题命中 效率与部署 :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05505 2025-09-09 cs.CL cs.LG 62%

Biomedical Literature Q&A System Using Retrieval-Augmented Generation (RAG)

Mansi Garg, Lee-Chi Wang, Bhavesh Ghanchi, Sanjana Dumpala, Shreyash Kakde, Yen Chih Chen

专题命中 效率与部署 :language model(abstract);分类 cs.CL、cs.LG

Comments 10 pages, 6 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00457 2025-09-09 cs.CL cs.LG 62%

CVPD at QIAS 2025 Shared Task: An Efficient Encoder-Based Approach for Islamic Inheritance Reasoning

Salah Eddine Bekhouche, Abdellah Zakaria Sellam, Hichem Telli, Cosimo Distante, Abdenour Hadid

机构 * University of the Basque Country UPV/EHU(巴斯克大学UPV/EHU) Institute of Applied Sciences and Intelligent Systems – CNR(应用科学与智能系统研究所 – CNR) Laboratory of LESIA, University of Biskra(LESIA实验室,比斯克拉大学) Sorbonne University Abu Dhabi, UAE(阿布扎赫尔索邦大学)

专题命中 效率与部署 :LLM(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08040 2025-09-09 cs.LG cs.AI 62%

BadPromptFL: A Novel Backdoor Threat to Prompt-based Federated Learning in Multimodal Models

Maozhen Zhang, Mengnan Zhao, Wei Wang, Bo Wang

机构 * School of Information and Communication Engineering, Dalian University of Technology(信息与通信工程学院,大连理工大学) School of Computer Science and Technology, Anhui University(计算机科学与技术学院,安徽大学) New Laboratory of Pattern Recognition (NLPR) State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS) Institute of Automation, Chinese Academy of Sciences (CASIA)(模式识别新实验室(NLPR)多模态人工智能系统国家重点实验室(MAIS)自动化研究所,中国科学院(CASIA))

专题命中 效率与部署 :language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06885 2025-09-09 cs.CV cs.AI 57%

Barlow-Swin: Toward a novel siamese-based segmentation architecture using Swin-Transformers

Morteza Kiani Haftlang, Mohammadhossein Malmir, Foroutan Parand, Umberto Michelucci, Safouane El Ghazouali

机构 * HSLU, Lucerne University of Applied Sciences and Arts(卢塞恩应用科学与艺术大学) Technical University of Munich(慕尼黑技术大学) University College London (UCL)(伦敦大学学院) TOELT LLC AI lab(TOELT LLC人工智能实验室)

专题命中 效率与部署 :pretraining(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05930 2025-09-09 cs.LG cs.SY eess.SY stat.ML 57%

Smoothed Online Optimization for Target Tracking: Robust and Learning-Augmented Algorithms

Ali Zeynali, Mahsa Sahebdel, Qingsong Liu, Mohammad Hajiesmaili, Ramesh K. Sitaraman

机构 * University of Massachusetts Amherst(马萨诸塞大学阿姆赫斯特分校) Akamai Technologies(Akamai技术公司)

专题命中 效率与部署 :LLM(abstract);分类 cs.LG

Comments 10 pages, 14 pages appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06220 2025-09-09 cs.CV cs.LG 57%

StreamMind: Unlocking Full Frame Rate Streaming Video Dialogue through Event-Gated Cognition

Xin Ding, Hao Wu, Yifan Yang, Shiqi Jiang, Donglin Bai, Zhibo Chen, Ting Cao

机构 * University of Science and Technology of China(中国科学技术大学) Microsoft Research(微软研究院) Nanjing University(南京大学) Institute for AI Industry Research (AIR), Tsinghua University(清华大学人工智能产业研究院)

专题命中 效率与部署 :LLM(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05651 2025-09-09 cs.MA cs.AI 57%

Orchestrator: Active Inference for Multi-Agent Systems in Long-Horizon Tasks

Lukas Beckenbauer, Johannes-Lucas Loewe, Ge Zheng, Alexandra Brintrup

机构 * Department of Engineering, University of Cambridge(剑桥大学工程系) TUM School of Management, Technical University of Munich(慕尼黑技术大学管理学院)

专题命中 效率与部署 :LLM(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18799 2025-09-09 cs.CV 50%

Robust and Label-Efficient Deep Waste Detection

Hassan Abid, Khan Muhammad, Muhammad Haris Khan

机构 * Mohamed Bin Zayed University of Artificial Intelligence(莫扎德·本·扎耶德人工智能大学) Sungkyunkwan University(成均馆大学)

专题命中 效率与部署 :LLM(abstract)

Comments Accepted at BMVC 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.09423 2025-09-09 cs.RO cs.CV 50%

Efficient Alignment of Unconditioned Action Prior for Language-conditioned Pick and Place in Clutter

Kechun Xu, Xunlong Xia, Kaixuan Wang, Yifei Yang, Yunxuan Mao, Bing Deng, Jieping Ye, Rong Xiong, Yue Wang

机构 * Zhejiang University and Alibaba Cloud(浙江大学和阿里云) Alibaba Cloud(阿里云) Zhejiang University(浙江大学)

专题命中 效率与部署 :foundation model(abstract)

Comments Accepted by T-ASE and CoRL25 GenPriors Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 领域大模型 28 篇

2505.07968 2025-09-09 cs.CL 90%

Assessing and Mitigating Medical Knowledge Drift and Conflicts in Large Language Models

Weiyi Wu, Xinwen Xu, Chongyang Gao, Xingjian Diao, Siting Li, Lucas A. Salas, Jiang Gui

机构 * Dartmouth College(达特茅斯学院) Massachusetts General Hospital(麻省总医院) Northwestern University(西北大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);preference optimization(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05390 2025-09-09 cs.CY cs.AI cs.CL 90%

Authorship Without Writing: Large Language Models and the Senior Author Analogy

Clint Hurshman, Sebastian Porsdam Mann, Julian Savulescu, Brian D. Earp

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

Comments 28 pages, 0 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.07825 2025-09-09 cs.CL 89%

Extracting and Combining Abilities For Building Multi-lingual Ability-enhanced Large Language Models

Zhipeng Chen, Kun Zhou, Liang Song, Wayne Xin Zhao, Bingning Wang, Weipeng Chen, Ji-Rong Wen

机构 * Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学人工智能学院) School of Information, Renmin University of China(中国人民大学信息学院) Baichuan Inc(百川科技)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

Comments EMNLP 2025 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22370 2025-09-09 cs.SE cs.PL 89%

Can Large Language Models Help Students Prove Software Correctness? An Experimental Study with Dafny

Carolina Carreira, Álvaro Silva, Alexandre Abreu, Alexandra Mendes

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05564 2025-09-09 cs.IR 89%

Knowledge-Augmented Relation Learning for Complementary Recommendation with Large Language Models

Chihiro Yamasaki, Kai Sugahara, Kazushi Okamoto

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract)

Journal ref The 2nd Workshop on Generative AI for E-Commerce 2025 in conjunction with the 19th ACM Conference on Recommender Systems (RecSys 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.00098 2025-09-09 cs.CL 88%

Fine-Tuning Large Language Models for Scientific Text Classification: A Comparative Study

Zhyar Rzgar K Rostam, Gábor Kertész

机构 * Doctoral School of Applied Informatics(应用信息学博士学院) Applied Mathematics Óbuda University Budapest, Hungary(应用数学 奥布达大学 布达佩斯 匈牙利) John von Neumann Faculty of Informatics(约翰·冯·诺依曼信息学学院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

Comments 6 pages, 3 figures, 7 tables

详情

展开后加载摘要…

URL PDF HTML 收藏