arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-10-16 至 2025-10-16 共收录 12 信号源:cs.CL, cs.AI, cs.LG

1. 预训练与数据 12 篇

2502.11671 2025-10-16 cs.CL cs.AI cs.LG 88%

Diversity-oriented Data Augmentation with Large Language Models

Zaitian Wang, Jinghan Zhang, Xinhao Zhang, Kunpeng Liu, Pengfei Wang, Yuanchun Zhou

机构 * Computer Network Information Center, CAS(中国科学院计算机网络信息中心) University of Chinese Academy of Sciences(中国科学院大学) Portland State University(波特兰州立大学)

专题命中 预训练与数据 :large language model(title);language model(title);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted to ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13697 2025-10-16 cs.SE cs.LG 85%

On Pretraining for Project-Level Code Completion

Maksim Sapronov, Evgeniy Glukhov

机构 * JetBrains Research(JetBrains 研究)

专题命中 预训练与数据 :pretraining(title,abstract);large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18185 2025-10-16 eess.SP cs.LG 83%

BrainOmni: A Brain Foundation Model for Unified EEG and MEG Signals

Qinfan Xiao, Ziyun Cui, Chi Zhang, Siqi Chen, Wen Wu, Andrew Thwaites, Alexandra Woolgar, Bowen Zhou, Chao Zhang

机构 * Shanghai Artificial Intelligence Laboratory, China(上海人工智能实验室) Department of Electronic Engineering, Tsinghua University, China(清华大学电子工程系) Department of Psychology, University of Cambridge, UK(剑桥大学心理学系) Speech Hearing and Phonetic Sciences, University College London, UK(伦敦大学学院语音听力与语音科学系) MRC Cognition and Brain Sciences Unit, University of Cambridge, UK(剑桥大学MRC认知与脑科学单位)

专题命中 预训练与数据 :foundation model(title,abstract);pretraining(abstract);分类 cs.LG

Comments Accepted by the 39th Conference on Neural Information Processing Systems (NeurIPS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.02340 2025-10-16 cs.CL cs.LG 82%

Can Prompts Rewind Time for LLMs? Evaluating the Effectiveness of Prompted Knowledge Cutoffs

Xin Gao, Ruiyi Zhang, Daniel Du, Saurabh Mahindre, Sai Ashish Somayajula, Pengtao Xie

机构 * UC San Diego(UC圣迭戈大学) SUNY Buffalo(纽约州立大学布法罗分校)

专题命中 预训练与数据 :large language model(abstract);language model(abstract);pretraining(abstract);prompting(abstract)

Comments Published at EMNLP 2025; Code and data available at https://github.com/gxx27/time_unlearn

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13462 2025-10-16 cs.CR 75%

Who Speaks for the Trigger? Dynamic Expert Routing in Backdoored Mixture-of-Experts Transformers

Xin Zhao, Xiaojun Chen, Bingshan Liu, Haoyu Gao, Zhendong Zhao, Yilong Chen

专题命中 预训练与数据 :large language model(abstract);language model(abstract);pretraining(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09710 2025-10-16 cs.CL cs.AI 73%

SeCon-RAG: A Two-Stage Semantic Filtering and Conflict-Free Framework for Trustworthy RAG

Xiaonan Si, Meilin Zhu, Simeng Qin, Lijia Yu, Lijun Zhang, Shuaitong Liu, Xinfeng Li, Ranjie Duan, Yang Liu, Xiaojun Jia

机构 * Institute of Software Chinese Academy of Sciences Beijing China(中国科学院软件研究所) Key Laboratory of System Software (Chinese Academy of Sciences) and State Key Laboratory of Computer Science, Institute of Software, Chinese Academy of Sciences, Beijing, China(中国科学院系统软件重点实验室和计算机科学国家重点实验室) University of Chinese Academy of Sciences, Beijing, China(中国科学院大学) Northeast University China(东北大学) Institute of Ai For industries Nanjing China(人工智能产业研究院) Southwest University China(西南大学) Nanyang Technological University Singapore(新加坡南洋理工大学) Alibaba China(阿里巴巴(中国))

专题命中 预训练与数据 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accepted at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12931 2025-10-16 cs.CV cs.CL 70%

Unifying Vision-Language Latents for Zero-label Image Caption Enhancement

Sanghyun Byun, Jung Ick Guack, Mohanad Odema, Baisub Lee, Jacob Song, Woo Seong Chung

机构 * LG Electronics USA(LG电子美国公司)

专题命中 预训练与数据 :language model(abstract);pretraining(abstract);分类 cs.CL

Comments Accepted to PMLR and NeurIPS 2025 UniReps

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.13211 2025-10-16 cs.CV cs.AI 70%

MIRROR: Multimodal Cognitive Reframing Therapy for Rolling with Resistance

Subin Kim, Hoonrae Kim, Jihyun Lee, Yejin Jeon, Gary Geunbae Lee

机构 * KT Corporation, Republic of Korea(韩国KT公司) Graduate School of Artificial Intelligence, POSTECH, Republic of Korea(POSTECH人工智能研究生院) Computer Science and Engineering, POSTECH, Republic of Korea(POSTECH计算机科学与工程系)

专题命中 预训练与数据 :large language model(abstract);language model(abstract);分类 cs.AI

Comments EMNLP 2025 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.03052 2025-10-16 cs.CL cs.LG 62%

Teaching Models to Understand (but not Generate) High-risk Data

Ryan Wang, Matthew Finlayson, Luca Soldaini, Swabha Swayamdipta, Robin Jia

机构 * Department of Computer Science, University of Southern California(计算机科学系,南加州大学) Allen Institute for AI(人工智能研究所)

专题命中 预训练与数据 :language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13768 2025-10-16 cs.CV cs.AI q-bio.NC 61%

Scaling Vision Transformers for Functional MRI with Flat Maps

Connor Lane, Daniel Z. Kaplan, Tanishq Mathew Abraham, Paul S. Scotti

机构 * Baylor College of Medicine(贝勒医学院) University of Florida(佛罗里达大学)

专题命中 预训练与数据 :foundation model(abstract,comments);分类 cs.AI

Comments NeurIPS 2025 Workshop, Foundation Models for the Brain and Body; Code: https://github.com/MedARC-AI/fmri-fm; Discord: https://discord.gg/tVR4TWnRM9

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13308 2025-10-16 eess.AS 50%

Towards Multimodal Query-Based Spatial Audio Source Extraction

Chenxin Yu, Hao Ma, Xu Li, Xiao-Lei Zhang, Mingjie Shao, Chi Zhang, Xuelong Li

专题命中 预训练与数据 :pretraining(abstract)

Comments Submitted to ICASSP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.16776 2025-10-16 cs.CV 50%

SynDiff-AD: Improving Semantic Segmentation and End-to-End Autonomous Driving with Synthetic Data from Latent Diffusion Models

Harsh Goel, Sai Shankar Narasimhan, Oguzhan Akcin, Sandeep Chinchali

机构 * The University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 预训练与数据 :prompting(abstract)

Comments 15 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏