arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-08-29 至 2025-08-29 共收录 117 信号源:cs.CL, cs.AI, cs.LG

1. 效率与部署 17 篇

2508.20914 2025-08-29 cs.SD cs.LG eess.AS 57%

Learning Robust Spatial Representations from Binaural Audio through Feature Distillation

Holger Severin Bovbjerg, Jan Østergaard, Jesper Jensen, Shinji Watanabe, Zheng-Hua Tan

机构 * Aalborg University(奥尔堡大学) Eriksholm Research Centre(埃里克肖尔研究中心) Carnegie Mellon University(卡内基梅隆大学)

专题命中 效率与部署 :pretraining(abstract);分类 cs.LG

Comments To appear in Proc. WASPAA 2025, October 12-15, 2025, Tahoe, US. Copyright (c) 2025 IEEE. 5 pages, 2 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20559 2025-08-29 cs.CL cs.IR 57%

Leveraging Generative Models for Real-Time Query-Driven Text Summarization in Large-Scale Web Search

Zeyu Xiong, Yixuan Nan, Li Gao, Hengzhu Tang, Shuaiqiang Wang, Junfeng Wang, Dawei Yin

机构 * Baidu Inc.(百度公司) Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)

专题命中 效率与部署 :preference optimization(abstract);分类 cs.CL

Comments CIKM'25

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 领域大模型 9 篇

2508.20217 2025-08-29 cs.CL cs.AI 91%

Prompting Strategies for Language Model-Based Item Generation in K-12 Education: Bridging the Gap Between Small and Large Language Models

Mohammad Amini, Babak Ahmadi, Xiaomeng Xiong, Yilin Zhang, Christopher Qiao

专题命中 领域大模型 :language model(title,abstract);prompting(title,abstract);large language model(title);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20453 2025-08-29 cs.CL 83%

MCP-Bench: Benchmarking Tool-Using LLM Agents with Complex Real-World Tasks via MCP Servers

Zhenting Wang, Qi Chang, Hemani Patel, Shashank Biju, Cheng-En Wu, Quan Liu, Aolin Ding, Alireza Rezazadeh, Ankit Shah, Yujia Bao, Eugene Siow

专题命中 领域大模型 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20437 2025-08-29 cs.LG cs.AI 81%

On Identifying Why and When Foundation Models Perform Well on Time-Series Forecasting Using Automated Explanations and Rating

Michael Widener, Kausik Lakkaraju, John Aydin, Biplav Srivastava

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.AI、cs.LG

Comments 8 pages, 5 Tables, 5 Figures, AI Trustworthiness and Risk Assessment for Challenged Contexts (ATRACC), Appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20345 2025-08-29 cs.CV cs.HC 78%

MedFoundationHub: A Lightweight and Secure Toolkit for Deploying Medical Vision Language Foundation Models

Xiao Li, Yanfan Zhu, Ruining Deng, Wei-Qi Wei, Yu Wang, Shilin Zhao, Yaohong Wang, Haichun Yang, Yuankai Huo

机构 * Vanderbilt University(范德比尔特大学) Weill Cornell Medicine(韦尔·科恩医学中心) Vanderbilt University Medical Center(范德比尔特大学医学中心) UT MD Anderson Cancer Center(德克萨斯大学MD安德森癌症中心)

专题命中 领域大模型 :foundation model(title);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20587 2025-08-29 cs.IR cs.LG 77%

SemSR: Semantics aware robust Session-based Recommendations

Jyoti Narwariya, Priyanka Gupta, Muskan Gupta, Jyotsana Khatri, Lovekesh Vig

机构 * TCS Research(TCS研究)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

Comments Accepted at EARL workshop @RecSys'25, Prague, Czech Republic

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01970 2025-08-29 cs.LG 77%

Improving Hospital Risk Prediction with Knowledge-Augmented Multimodal EHR Modeling

Rituparna Datta, Jiaming Cui, Zihan Guan, Vishal G. Reddy, Joshua C. Eby, Gregory Madden, Rupesh Silwal, Anil Vullikanti

机构 * Department of Computer Science, University of Virginia(大学计算机科学系) University of Virginia School of Medicine(弗吉尼亚大学医学院) Virginia Polytechnic Institute and State University(弗吉尼亚理工学院和州立大学) Biocomplexity Institute and Initiative, University of Virginia(大学生物复杂性研究所) Division of Infectious Diseases & International Health, University of Virginia School of Medicine(大学感染性疾病与国际卫生分会)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.21024 2025-08-29 cs.CL cs.IR 70%

An Agile Method for Implementing Retrieval Augmented Generation Tools in Industrial SMEs

Mathieu Bourdin, Anas Neumann, Thomas Paviot, Robert Pellerin, Samir Lamouri

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

Comments 20 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.14330 2025-08-29 cs.SE 67%

Leveraging LLMs for Formal Software Requirements -- Challenges and Prospects

Arshad Beg, Diarmuid O'Donoghue, Rosemary Monahan

专题命中 领域大模型 :large language model(abstract);language model(abstract)

Comments Overlay2025 - 7th International Workshop on Artificial Intelligence and fOrmal VERification, Logic, Automata, and sYnthesis. [Accepted]. To be held on 26th of October, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20139 2025-08-29 eess.IV cs.CV cs.HC cs.LG 57%

Is the medical image segmentation problem solved? A survey of current developments and future directions

Guoping Xu, Jayaram K. Udupa, Jax Luo, Songlin Zhao, Yajun Yu, Scott B. Raymond, Hao Peng, Lipeng Ning, Yogesh Rathi, Wei Liu, You Zhang

专题命中 领域大模型 :foundation model(abstract);分类 cs.LG

Comments 80 pages, 38 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 知识编辑与模型理解 5 篇

2508.20805 2025-08-29 cs.CL cs.AI cs.SD 84%

Exploring Machine Learning and Language Models for Multimodal Depression Detection

Javier Si Zhao Hong, Timothy Zoe Delaya, Sherwyn Chan Yin Kit, Pai Chet Ng, Xiaoxiao Miao

机构 * Singapore Institute of Technology(新加坡理工学院) Duke Kunshan University(杜克-昆山大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI

Comments This paper has been accepted by APCIPA ASC 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.13722 2025-08-29 cs.CL 79%

Probing Pre-Trained Language Models for Cross-Cultural Differences in Values

Arnav Arora, Lucie-Aimée Kaffee, Isabelle Augenstein

机构 * University of Copenhagen(哥本哈根大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments Accepted to C3NLP, EACL 2023: https://aclanthology.org/2023.c3nlp-1.12/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20578 2025-08-29 cs.AI cs.CR 77%

Human-AI Collaborative Bot Detection in MMORPGs

Jaeman Son, Hyunsoo Kim

机构 * NCSOFT

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20195 2025-08-29 cs.AI cs.CL cs.MA 73%

AI-AI Esthetic Collaboration with Explicit Semiotic Awareness and Emergent Grammar Development

Nicanor I. Moldovan

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 13 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05200 2025-08-29 cs.LG math.ST stat.ML stat.TH 70%

Transformers Meet In-Context Learning: A Universal Approximation Theory

Gen Li, Yuchen Jiao, Yu Huang, Yuting Wei, Yuxin Chen

机构 * Department of Statistics and Data Science, Chinese University of Hong Kong(统计与数据科学系,香港中文大学) the Wharton School, University of Pennsylvania(宾夕法尼亚大学沃顿商学院) University of Pennsylvania(宾夕法尼亚大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 其他LLM 11 篇

2508.02997 2025-08-29 cs.CL 88%

CoCoTen: Detecting Adversarial Inputs to Large Language Models through Latent Space Features of Contextual Co-occurrence Tensors

Sri Durga Sai Sowmya Kadali, Evangelos E. Papalexakis

机构 * University of California, Riverside(加州大学河滨分校)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20577 2025-08-29 cs.LG cs.AI cs.CL 85%

MERIT: Maximum-normalized Element-wise Ratio for Language Model Large-batch Training

Yang Luo, Zangwei Zheng, Ziheng Qin, Zirui Zhu, Yong Liu, Yang You

机构 * School of Computing, National University of Singapore, Singapore(新加坡国立大学计算机学院)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20263 2025-08-29 cs.HC 85%

Athena: Intermediate Representations for Iterative Scaffolded App Generation with an LLM

Jazbo Beason, Ruijia Cheng, Eldon Schoop, Jeffrey Nichols

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20186 2025-08-29 cs.CR cs.AI cs.CY 83%

AI Propaganda factories with language models

Lukasz Olejnik

机构 * Department of War Studies(战争研究系) King's College London(伦敦国王学院)

专题命中 其他LLM :language model(title,abstract);small language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20333 2025-08-29 cs.LG cs.AI cs.CL cs.DC 80%

Poison Once, Refuse Forever: Weaponizing Alignment for Injecting Bias in LLMs

Md Abdullah Al Mamun, Ihsen Alouani, Nael Abu-Ghazaleh

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.15692 2025-08-29 cs.LG 77%

MLE-STAR: Machine Learning Engineering Agent via Search and Targeted Refinement

Jaehyun Nam, Jinsung Yoon, Jiefeng Chen, Jinwoo Shin, Sercan Ö. Arık, Tomas Pfister

机构 * Google Cloud(谷歌云) KAIST(韩国科学技术院)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10439 2025-08-29 cs.CV cs.RO 75%

CogNav: Cognitive Process Modeling for Object Goal Navigation with LLMs

Yihan Cao, Jiazhao Zhang, Zhinan Yu, Shuzhen Liu, Zheng Qin, Qin Zou, Bo Du, Kai Xu

机构 * College of Computer Science and Technology, National University of Defense Technology(计算机科学与技术学院,国防科技大学) CFCS, School of Computer Science, Peking University(计算机科学系,北京大学) Defense Innovation Institute, Academy of Military Sciences(国防科技创新研究院,军事科学院) School of Computer Science, Wuhan University(计算机科学学院,武汉大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20635 2025-08-29 cs.HC 67%

Schema-Guided Response Generation using Multi-Frame Dialogue State for Motivational Interviewing Systems

Jie Zeng, Yukiko I. Nakano

专题命中 其他LLM :large language model(abstract);language model(abstract)

Comments 28pages, 15 figures, 10 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06989 2025-08-29 cs.CR cs.CV 67%

Probabilistic Modeling of Jailbreak on Multimodal LLMs: From Quantification to Application

Wenzhuo Xu, Zhipeng Wei, Xiongtao Sun, Zonghao Ying, Deyue Zhang, Dongdong Yang, Xiangzheng Zhang, Quanchen Zou

专题命中 其他LLM :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.18674 2025-08-29 cs.CV 67%

Image-guided topic modeling for interpretable privacy classification

Alina Elena Baia, Andrea Cavallaro

专题命中 其他LLM :large language model(abstract);language model(abstract)

Comments Paper accepted at the eXCV Workshop at ECCV 2024. Supplementary material included. Code available at https://github.com/idiap/itm

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.14316 2025-08-29 cs.CV 50%

T-Stars-Poster: A Framework for Product-Centric Advertising Image Design

Hongyu Chen, Min Zhou, Jing Jiang, Jiale Chen, Yang Lu, Zihang Lin, Bo Xiao, Tiezheng Ge, Bo Zheng

机构 * Alibaba Group(阿里巴巴集团) Beijing University of Posts and Telecommunications(北京邮电大学)

专题命中 其他LLM :language model(abstract)

Comments Accepted by CIKM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏