arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-07-31 至 2025-07-31 共收录 120 信号源:cs.CL, cs.AI, cs.LG

1. 预训练与数据 13 篇

2507.22187 2025-07-31 cs.CL cs.AI cs.LG 90%

A Scalable Pipeline for Estimating Verb Frame Frequencies Using Large Language Models

Adam M. Morgan, Adeen Flinker

机构 * NYU Grossman School of Medicine(纽约大学格罗斯曼医学院) NYU Tandon School of Engineering(纽约大学坦顿工程学院)

专题命中 预训练与数据 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19703 2025-07-31 cs.AI 89%

The wall confronting large language models

Peter V. Coveney, Sauro Succi

机构 * Centre for Computational Science, University College London(伦敦大学学院计算科学中心) Advanced Research Computing Centre, University College London(伦敦大学学院高级计算中心) Institute for Informatics, Faculty of Science, University of Amsterdam(阿姆斯特丹大学科学学院信息学院) Italian Institute of Technology(意大利理工学院) Physics Department, Harvard University(哈佛大学物理系) Mechanical Engineering Department, University College London(伦敦大学学院机械工程系)

专题命中 预训练与数据 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22514 2025-07-31 cs.LG 88%

SmilesT5: Domain-specific pretraining for molecular language models

Philip Spence, Brooks Paige, Anne Osbourn

专题命中 预训练与数据 :language model(title,abstract);pretraining(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.14373 2025-07-31 cs.CL eess.SP 87%

ECG-Byte: A Tokenizer for End-to-End Generative Electrocardiogram Language Modeling

William Han, Chaojing Duan, Michael A. Rosenberg, Emerson Liu, Ding Zhao

机构 * Carnegie Mellon University(卡内基梅隆大学) Allegheny Health Network(阿勒格尼健康网络) University of Colorado(科罗拉多大学)

专题命中 预训练与数据 :language model(title,abstract);LLM(abstract);large language model(abstract);pretraining(abstract)

Comments 38 pages, 9 figures; Accepted to MLHC 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.06157 2025-07-31 cs.SI cs.CL 85%

Masked Language Models are Good Heterogeneous Graph Generalizers

Jinyu Yang, Cheng Yang, Shanyuan Cui, Zeyuan Guo, Liangwei Yang, Muhan Zhang, Zhiqiang Zhang, Chuan Shi

机构 * School of Computer Science, Beijing University of Posts and Telecommunications(北京邮电大学计算机科学学院) Department of Computer Science, University of Illinois Chicago(伊利诺伊大学芝加哥分校计算机科学系) Institute for Artificial Intelligence, Peking University(北京大学人工智能研究院) Ant Group, Beijing, China(蚂蚁集团(北京))

专题命中 预训练与数据 :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22431 2025-07-31 cs.CV 82%

HQ-CLIP: Leveraging Large Vision-Language Models to Create High-Quality Image-Text Datasets and CLIP Models

Zhixiang Wei, Guangting Wang, Xiaoxiao Ma, Ke Mei, Huaian Chen, Yi Jin, Fengyun Rao

机构 * University of Science and Technology of China(中国科学技术大学) WeChat Vision, Tencent Inc.(腾讯公司)

专题命中 预训练与数据 :language model(title,abstract);pretraining(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22210 2025-07-31 q-bio.QM 82%

Scaling and Data Saturation in Protein Language Models

Aviv Spinner, Erika DeBenedictis, Corey M. Hudson

专题命中 预训练与数据 :language model(title,abstract);pretraining(abstract)

Comments Presented at the GenBio Workshop, ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.13540 2025-07-31 cs.LG 77%

Provable Low-Frequency Bias of In-Context Learning of Representations

Yongyi Yang, Hidenori Tanaka, Wei Hu

机构 * Computer Science and Engineering, University of Michigan(密歇根大学计算机科学与工程系) CBS-NTT Physics of Intelligence Program, Harvard University(哈佛大学CBS-NTT智能物理计划) Physics of Artificial Intelligence Group, NTT Research, Inc.(NTT研究公司人工智能物理小组)

专题命中 预训练与数据 :large language model(abstract);language model(abstract);pretraining(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.08549 2025-07-31 cs.SE 75%

Vulnerability Handling of AI-Generated Code -- Existing Solutions and Open Challenges

Sabrina Kaniewski, Dieter Holstein, Fabian Schmidt, Tobias Heer

专题命中 预训练与数据 :LLM(abstract);large language model(abstract);language model(abstract)

Comments Accepted for publication @ IEEE AIxSET 2024; 4 pages, 2 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22250 2025-07-31 cs.LG cs.AI 73%

Using Scaling Laws for Data Source Utility Estimation in Domain-Specific Pre-Training

Oleksiy Ostapenko, Charles Guille-Escuret, Luke Kumar, Max Tian, Denis Kocetkov, Gopeshh Subbaraj, Raymond Li, Joel Lamy-Poirier, Sebastien Paquet, Torsten Scholak

机构 * ServiceNow Research(ServiceNow 研究所) Mila — Quebec AI Institute(魁北克人工智能研究所)

专题命中 预训练与数据 :foundation model(abstract);pretraining(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22606 2025-07-31 cs.AI 70%

MetaAgent: Automatically Constructing Multi-Agent Systems Based on Finite State Machines

Yaolun Zhang, Xiaogeng Liu, Chaowei Xiao

机构 * University of Wisconsin - Madison(威斯康星大学麦迪逊分校)

专题命中 预训练与数据 :large language model(abstract);language model(abstract);分类 cs.AI

Comments ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22404 2025-07-31 cs.CV cs.AI cs.LG 62%

MINR: Implicit Neural Representations with Masked Image Modelling

Sua Lee, Joonhun Lee, Myungjoo Kang

机构 * Seoul National University(首尔国立大学)

专题命中 预训练与数据 :pretraining(abstract);分类 cs.AI、cs.LG

Comments Accepted to the ICCV 2023 workshop on Out-of-Distribution Generalization in Computer Vision

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22378 2025-07-31 eess.IV cs.CV 50%

Whole-brain Transferable Representations from Large-Scale fMRI Data Improve Task-Evoked Brain Activity Decoding

Yueh-Po Peng, Vincent K. M. Cheung, Li Su

机构 * Gamania Digital Entertainment Co., Ltd.(Gamania数字娱乐有限公司) Institute of Information Science, Academia Sinica(Academia Sinica信息科学研究所) Sony Computer Science Laboratories, Inc.(索尼计算机科学实验室有限公司)

专题命中 预训练与数据 :pretraining(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 指令微调 8 篇

2412.01233 2025-07-31 cs.AI 90%

Best Practices for Large Language Models in Radiology

Christian Bluethgen, Dave Van Veen, Cyril Zakka, Katherine Link, Aaron Fanous, Roxana Daneshjou, Thomas Frauenfelder, Curtis Langlotz, Sergios Gatidis, Akshay Chaudhari

机构 * Stanford Center for Artificial Intelligence in Medicine and Imaging(斯坦福大学人工智能医学与成像中心) Diagnostic and Interventional Radiology, University Hospital Zurich, University of Zurich(苏黎世大学医院放射诊断与介入放射学) Department of Electrical Engineering, Stanford University(斯坦福大学电气工程系) Department of Cardiothoracic Surgery, Stanford Medicine(斯坦福医学院心胸外科系) Hugging Face Department of Medical Education, Icahn School of Medicine at Mount Sinai(伊坎医学院Mount Sinai医学教育系) NVIDIA Corporation(NVIDIA公司) UT Health San Antonio(UT健康科学中心圣安东尼奥分校) Department of Dermatology, Redwood City, CA, USA(红木城加州大学皮肤病学系) Department of Biomedical Data Science, Stanford, CA, USA(斯坦福大学生物医学数据科学系) Department of Medicine, Stanford, CA, USA(斯坦福大学医学系) Department of Radiology, Stanford University(斯坦福大学放射学系)

专题命中 指令微调 :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

Comments A redacted version of this preprint has been accepted for publication in Radiology

Journal ref Radiology 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.08003 2025-07-31 cs.AI cs.CL cs.CY cs.FL 90%

Can adversarial attacks by large language models be attributed?

Manuel Cebrian, Andres Abeliuk, Jan Arne Telle

机构 * Center for Automation and Robotics, Spanish National Research Council(自动化与机器人中心,西班牙国家研究理事会) Department of Computer Science, University of Chile(计算机科学系,智利大学) Department of Informatics, University of Bergen(信息学系,卑尔根大学)

专题命中 指令微调 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

Comments 22 pages, 5 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.09213 2025-07-31 cs.CL 87%

FineMedLM-o1: Enhancing Medical Knowledge Reasoning Ability of LLM from Supervised Fine-Tuning to Test-Time Training

Hongzhou Yu, Tianhao Cheng, Yingwen Wang, Wen He, Qing Wang, Ying Cheng, Yuejie Zhang, Rui Feng, Xiaobo Zhang

机构 * Fudan University(复旦大学) Children’s Hospital of Fudan University(复旦大学附属儿童医院)

专题命中 指令微调 :LLM(title);large language model(abstract);language model(abstract);SFT(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22469 2025-07-31 cs.CV cs.AI cs.LG 82%

Visual Language Models as Zero-Shot Deepfake Detectors

Viacheslav Pirogov

专题命中 指令微调 :language model(title,abstract);分类 cs.AI、cs.LG;foundation model(comments)

Comments Accepted to the ICML 2025 Workshop on Reliable and Responsible Foundation Models

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.00640 2025-07-31 cs.AI 77%

CollabLLM: From Passive Responders to Active Collaborators

Shirley Wu, Michel Galley, Baolin Peng, Hao Cheng, Gavin Li, Yao Dou, Weixin Cai, James Zou, Jure Leskovec, Jianfeng Gao

机构 * stanford(斯坦福大学) microsoft(微软公司)

专题命中 指令微调 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments Outstanding Paper Award at ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04858 2025-07-31 cs.AI cs.LG 62%

Don't Lag, RAG: Training-Free Adversarial Detection Using RAG

Roie Kazoom, Raz Lapid, Moshe Sipper, Ofer Hadar

机构 * Electrical and Computer Engineering, Ben Gurion University, Beer Sheba 84105, Israel(电子与计算机工程系,本· Gurion 大学) Computer Science, Ben Gurion University, Beer Sheba 84105, Israel(计算机科学系,本· Gurion 大学)

专题命中 指令微调 :language model(abstract);分类 cs.AI、cs.LG

Comments Accepted at VecDB @ ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04395 2025-07-31 cs.LG 57%

Human-Level Competitive Pokémon via Scalable Offline Reinforcement Learning with Transformers

Jake Grigsby, Yuqi Xie, Justin Sasek, Steven Zheng, Yuke Zhu

专题命中 指令微调 :LLM(abstract);分类 cs.LG

Comments Reinforcement Learning Conference 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22003 2025-07-31 cs.CV 50%

See Different, Think Better: Visual Variations Mitigating Hallucinations in LVLMs

Ziyun Dai, Xiaoqiang Li, Shaohua Zhang, Yuanchen Wu, Jide Li

专题命中 指令微调 :language model(abstract)

Comments Accepted by ACM MM25

Journal ref 33rd ACM International Conference on Multimedia 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 后训练与偏好优化 2 篇

2507.21391 2025-07-31 cs.CV cs.AI cs.CL 73%

Multimodal LLMs as Customized Reward Models for Text-to-Image Generation

Shijie Zhou, Ruiyi Zhang, Huaisheng Zhu, Branislav Kveton, Yufan Zhou, Jiuxiang Gu, Jian Chen, Changyou Chen

机构 * University at Buffalo(布法罗大学) Adobe Research(Adobe研究) Pennsylvania State University(宾夕法尼亚州立大学)

专题命中 后训练与偏好优化 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accepted at ICCV 2025. Code available at https://github.com/sjz5202/LLaVA-Reward

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22356 2025-07-31 cs.RO 50%

In-Situ Soil-Property Estimation and Bayesian Mapping with a Simulated Compact Track Loader

W. Jacob Wagner, Ahmet Soylemezoglu, Katherine Driggs-Campbell

机构 * W. Jacob Wagner(独立研究者) Ahmet Soylemezoglu(独立研究者) Katherine Driggs-Campbell(独立研究者)

专题命中 后训练与偏好优化 :post-training(abstract)

Comments 29 pages, 12 figures, 5 algorithms, ISTVS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 长上下文与记忆 1 篇

2504.14200 2025-07-31 cs.CV cs.AI 57%

Enhancing Multimodal In-Context Learning for Image Classification through Coreset Optimization

Huiyi Chen, Jiawei Peng, Kaihua Tang, Xin Geng, Xu Yang

机构 * Southeast University(东南大学) Huawei Singapore Research Center(华为新加坡研究中心)

专题命中 长上下文与记忆 :language model(abstract);分类 cs.AI

Comments 11 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 推理与问题求解 10 篇

2505.21354 2025-07-31 cs.CL cs.LG 90%

Leveraging Large Language Models for Bengali Math Word Problem Solving with Chain of Thought Reasoning

Bidyarthi Paul, Jalisha Jashim Era, Mirazur Rahman Zim, Tahmid Sattar Aothoi, Faisal Muhammad Shah

机构 * Ahsanullah University of Science and Technology(阿沙努拉科技大学)

专题命中 推理与问题求解 :large language model(title,abstract);language model(title,abstract);prompting(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22467 2025-07-31 cs.MA cs.AI cs.CY 85%

Towards Simulating Social Influence Dynamics with LLM-based Multi-agents

Hsien-Tsung Lin, Pei-Cing Huang, Chan-Tung Ku, Chan Hsu, Pei-Xuan Shieh, Yihuang Kang

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.15790 2025-07-31 cs.CR cs.SE 85%

ETrace:Event-Driven Vulnerability Detection in Smart Contracts via LLM-Based Trace Analysis

Chenyang Peng, Haijun Wang, Yin Wu, Hao Wu, Ming Fan, Yitao Zhao, Ting Liu

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract)

Comments 4 pages, 1 figure. To appear in Proceedings of the 16th Asia-Pacific Symposium on Internetware (Internetware 2025), ACM ICPS. DOI: https://doi.org/10.1145/3755881.3755934

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.13820 2025-07-31 cs.AI cs.CL cs.LG cs.SE 80%

Scoring Verifiers: Evaluating Synthetic Verification for Code and Reasoning

Aleksander Ficek, Somshubra Majumdar, Vahid Noroozi, Boris Ginsburg

机构 * NVIDIA Santa Clara, CA 15213, USA(英伟达圣克拉拉分校)

专题命中 推理与问题求解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments COLM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22887 2025-07-31 cs.CL cs.AI 79%

Where to show Demos in Your Prompt: A Positional Bias of In-Context Learning

Kwesi Cobbina, Tianyi Zhou

机构 * University of Maryland, College Park(马里兰大学 College Park分校)

专题命中 推理与问题求解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.15225 2025-07-31 cs.AI cs.LG cs.LO 79%

Solving Formal Math Problems by Decomposition and Iterative Reflection

Yichi Zhou, Jianqiu Zhao, Yongxin Zhang, Bohan Wang, Siran Wang, Luoxin Chen, Jiahui Wang, Haowei Chen, Allan Jie, Xinbo Zhang, Haocheng Wang, Luong Trung, Rong Ye, Phan Nhat Hoang, Huishuai Zhang, Peng Sun, Hang Li

机构 * University of science and technology of China(科学技术大学) Peking University(北京大学)

专题命中 推理与问题求解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏