arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-11-03 至 2025-11-03 共收录 160 信号源:cs.CL, cs.AI, cs.LG

1. 效率与部署 30 篇

2510.27401 2025-11-03 cs.HC 50%

"Koyi Sawaal Nahi Hai": Reimagining Maternal Health Chatbots for Collective, Culturally Grounded Care

Imaan Hameed, Huma Umar, Fozia Umber, Maryam Mustafa

专题命中 效率与部署 :LLM(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27135 2025-11-03 cs.CV 50%

E-MMDiT: Revisiting Multimodal Diffusion Transformer Design for Fast Image Synthesis under Limited Resources

Tong Shen, Jingai Yu, Dong Zhou, Dong Li, Emad Barsoum

机构 * Advanced Micro Devices, Inc.(Advanced Micro Devices公司)

专题命中 效率与部署 :post-training(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 领域大模型 18 篇

2510.27080 2025-11-03 cs.CR cs.AI 89%

Adapting Large Language Models to Emerging Cybersecurity using Retrieval Augmented Generation

Arnabh Borah, Md Tanvirul Alam, Nidhi Rastogi

机构 * School of Electrical and Computer Engineering(电气与计算机工程学院) Georgia Institute of Technology(佐治亚理工学院) Department of Computer Science(计算机科学系) Rochester Institute of Technology(罗切斯特理工学院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.00455 2025-11-03 cs.HC cs.AI 88%

Data Therapist: Eliciting Domain Knowledge from Subject Matter Experts Using Large Language Models

Sungbok Shin, Hyeon Jeon, Sanghyun Hong, Niklas Elmqvist

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.12170 2025-11-03 cs.NE 88%

Language Model Crossover: Variation through Few-Shot Prompting

Elliot Meyerson, Mark J. Nelson, Herbie Bradley, Adam Gaier, Arash Moradi, Amy K. Hoover, Joel Lehman

专题命中 领域大模型 :language model(title,abstract);prompting(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27628 2025-11-03 cs.AI 77%

Validity Is What You Need

Sebastian Benthall, Andrew Clark

机构 * International Computer Science Institute(国际计算机科学研究所) New York University School of Law(纽约大学法学院)

专题命中 领域大模型 :large language model(abstract);language model(abstract);foundation model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27535 2025-11-03 cs.CL 77%

Patient-Centered Summarization Framework for AI Clinical Summarization: A Mixed-Methods Design

Maria Lizarazo Jimenez, Ana Gabriela Claros, Kieran Green, David Toro-Tobon, Felipe Larios, Sheena Asthana, Camila Wenczenovicz, Kerly Guevara Maldonado, Luis Vilatuna-Andrango, Cristina Proano-Velez, Satya Sai Sri Bandi, Shubhangi Bagewadi, Megan E. Branda, Misk Al Zahidy, Saturnino Luz, Mirella Lapata, Juan P. Brito, Oscar J. Ponce-Ponte

专题命中 领域大模型 :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.CL

Comments The first two listed authors contributed equally Pages: 21; Figures:2; Tables:3

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26824 2025-11-03 cs.DL cs.AI cs.IR 77%

LeMat-Synth: a multi-modal toolbox to curate broad synthesis procedure databases from scientific literature

Magdalena Lederbauer, Siddharth Betala, Xiyao Li, Ayush Jain, Amine Sehaba, Georgia Channing, Grégoire Germain, Anamaria Leonescu, Faris Flaifil, Alfonso Amayuelas, Alexandre Nozadze, Stefan P. Schmid, Mohd Zaki, Sudheesh Kumar Ethirajan, Elton Pan, Mathilde Franckel, Alexandre Duval, N. M. Anoop Krishnan, Samuel P. Gleason

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments 29 pages, 13 figures, 6 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26996 2025-11-03 cs.CV 75%

MoME: Mixture of Visual Language Medical Experts for Medical Imaging Segmentation

Arghavan Rezvani, Xiangyi Yan, Anthony T. Wu, Kun Han, Pooya Khosravi, Xiaohui Xie

机构 * Department of Computer Science, University of California, Irvine(加州大学尔湾分校计算机科学系) School of Medicine, University of California, Irvine(加州大学尔湾分校医学院)

专题命中 领域大模型 :large language model(abstract);language model(abstract);foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27552 2025-11-03 cs.CL 74%

Multilingual BERT language model for medical tasks: Evaluation on domain-specific adaptation and cross-linguality

Yinghao Luo, Lang Zhou, Amrish Jhingoer, Klaske Vliegenthart Jongbloed, Carlijn Jordans, Ben Werkhoven, Tom Seinen, Erik van Mulligen, Casper Rokx, Yunlei Li

专题命中 领域大模型 :language model(title);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26974 2025-11-03 cs.CL cs.AI 73%

Overview of the MEDIQA-OE 2025 Shared Task on Medical Order Extraction from Doctor-Patient Consultations

Jean-Philippe Corbeil, Asma Ben Abacha, Jerome Tremblay, Phillip Swazinna, Akila Jeeson Daniel, Miguel Del-Agua, Francois Beaulieu

机构 * Microsoft Healthcare & Life Sciences(微软医疗与生命科学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27130 2025-11-03 cs.LG 70%

AI Agents in Drug Discovery

Srijit Seal, Dinh Long Huynh, Moudather Chelbi, Sara Khosravi, Ankur Kumar, Mattson Thieme, Isaac Wilks, Mark Davies, Jessica Mustali, Yannick Sun, Nick Edwards, Daniil Boiko, Andrei Tyrin, Douglas W. Selinger, Ayaan Parikh, Rahul Vijayan, Shoman Kasbekar, Dylan Reid, Andreas Bender, Ola Spjuth

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.LG

Comments 45 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.09298 2025-11-03 cs.LG 70%

DeepOSets: Non-Autoregressive In-Context Learning with Permutation-Invariance Inductive Bias

Shao-Ting Chiu, Junyuan Hong, Ulisses Braga-Neto

机构 * Dept of Electrical & Computer Engineering Texas A&M University College Station, TX, USA(电气与计算机工程系塔拉斯州立大学) Department of Electrical & Computer Engineering University of Texas at Austin(电气与计算机工程系德克萨斯大学奥斯汀分校)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.LG

Comments Set transformer results in the high-dimensional (d=20) case were added; there is a revised proof of Theorem 1; minor edits were made throughout

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26886 2025-11-03 cond-mat.mtrl-sci 67%

MaterialsGalaxy: A Platform Fusing Experimental and Theoretical Data in Condensed Matter Physics

Tiannian Zhu, Zhong Fang, Quansheng Wu, Hongming Weng

专题命中 领域大模型 :large language model(abstract);language model(abstract)

Comments 18 pages, 5 figures. Accepted for publication in Chinese Physics B (24 October 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23607 2025-11-03 cs.LG cs.AI cs.CL 67%

Deep Learning-based Prediction of Clinical Trial Enrollment with Uncertainty Estimates

Tien Huu Do, Antoine Masquelier, Nae Eoun Lee, Jonathan Crowther

机构 * Pfizer(辉瑞公司) Merck(默克公司)

专题命中 领域大模型 :language model(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27213 2025-11-03 cs.CV cs.AI cs.LG 62%

Privacy-Aware Continual Self-Supervised Learning on Multi-Window Chest Computed Tomography for Domain-Shift Robustness

Ren Tasai, Guang Li, Ren Togo, Takahiro Ogawa, Kenji Hirata, Minghui Tang, Takaaki Yoshimura, Hiroyuki Sugimori, Noriko Nishioka, Yukie Shimizu, Kohsuke Kudo, Miki Haseyama

机构 * Hokkaido University(北海道大学)

专题命中 领域大模型 :pretraining(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27265 2025-11-03 cs.CV cs.LG 57%

T3: Test-Time Model Merging in VLMs for Zero-Shot Medical Imaging Analysis

Raza Imam, Hu Wang, Dwarikanath Mahapatra, Mohammad Yaqub

机构 * Mohammed bin Zayed University of Artificial Intelligence(莫卧儿bin Zayed人工智能大学)

专题命中 领域大模型 :language model(abstract);分类 cs.LG

Comments Main: 11 pages, Supplementary: 9 pages 10 tables, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24204 2025-11-03 cs.CV cs.AI 57%

BALR-SAM: Boundary-Aware Low-Rank Adaptation of SAM for Resource-Efficient Medical Image Segmentation

Zelin Liu, Sicheng Dong, Bocheng Li, Yixuan Yang, Jiacheng Ruan, Chenxu Zhou, Suncheng Xiang

机构 * Shanghai Jiao Tong University(上海交通大学)

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18384 2025-11-03 cs.CR cs.AI 57%

Dynamic Risk Assessments for Offensive Cybersecurity Agents

Boyi Wei, Benedikt Stroebl, Jiacen Xu, Joie Zhang, Zhou Li, Peter Henderson

机构 * Princeton University(普林斯顿大学) Microsoft(微软公司) University of California Irvine(加州大学尔湾分校)

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI

Comments 26 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27452 2025-11-03 cs.CV 50%

From Pixels to Paths: A Multi-Agent Framework for Editable Scientific Illustration

Jianwen Sun, Fanrui Zhang, Yukang Feng, Chuanhao Li, Zizhen Li, Jiaxin Ai, Yifan Chang, Yu Dai, Kaipeng Zhang

机构 * Nankai University(南开大学) Shanghai Innovation Institute(上海创新研究院) Wuhan University(武汉大学) University of Science and Technology of China(中国科学技术大学) Shanghai AI Laboratory(上海人工智能实验室)

专题命中 领域大模型 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 知识编辑与模型理解 8 篇

2510.27355 2025-11-03 cs.CL 85%

ThoughtProbe: Classifier-Guided LLM Thought Space Exploration via Probing Representations

Zijian Wang, Chang Xu

机构 * School of Computer Science(计算机科学学院) The University of Sydney(悉尼大学)

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments EMNLP2025 main conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27131 2025-11-03 cs.LG 84%

Exploring the Utilities of the Rationales from Large Language Models to Enhance Automated Essay Scoring

Hong Jiao, Hanna Choi, Haowei Hua

专题命中 知识编辑与模型理解 :large language model(title);language model(title);分类 cs.LG

Comments 12 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20278 2025-11-03 q-bio.QM cs.LG 83%

The cell as a token: high-dimensional geometry in language models and cell embeddings

William Gilpin

机构 * Department of Physics, The University of Texas at Austin, Austin, Texas 78712, USA(德克萨斯大学奥斯汀分校物理系)

专题命中 知识编辑与模型理解 :language model(title);foundation model(abstract);pretraining(abstract);分类 cs.LG

Comments 4 pages, 2 figures

Journal ref Bioinformatics (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.16724 2025-11-03 cs.AI cs.LG 81%

Towards Automated Semantic Interpretability in Reinforcement Learning via Vision-Language Models

Zhaoxin Li, Zhang Xi-Jia, Batuhan Altundas, Letian Chen, Rohan Paleja, Matthew Gombolay

机构 * Georgia Institute of Technology(佐治亚理工学院) Purdue University(普渡大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.00883 2025-11-03 cs.CL 77%

Mathematics Isn't Culture-Free: Probing Cultural Gaps via Entity and Scenario Perturbations

Aditya Tomar, Nihar Ranjan Sahoo, Ashish Mittal, Rudra Murthy, Pushpak Bhattacharyya

机构 * IIT Bombay(印度班加罗尔理工学院) IBM Research(IBM研究院)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.00333 2025-11-03 cs.CL cs.AI 73%

More of the Same: Persistent Representational Harms Under Increased Representation

Jennifer Mickel, Maria De-Arteaga, Leqi Liu, Kevin Tian

机构 * EleutherAI Universitat Ramon Llull, ESADE(拉蒙·拉鲁尔大学,ESADE) UT Austin(德克萨斯大学奥斯汀分校)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Proceedings of the Neural Information Processing Systems (NeurIPS) 2025; 39 pages, 7 figures, 15 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11295 2025-11-03 cs.CV 67%

Human Uncertainty-Aware Data Selection and Automatic Labeling in Visual Question Answering

Jian Lan, Zhicheng Liu, Udo Schlegel, Raoyuan Zhao, Yihong Liu, Hinrich Schütze, Michael A. Hedderich, Thomas Seidl

机构 * University of Munich(慕尼黑大学) Munich Center of Machine Learning(慕尼黑机器学习中心)

专题命中 知识编辑与模型理解 :language model(abstract);SFT(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.23953 2025-11-03 cs.LG cs.AI cs.CL cs.CY cs.GT 67%

Representative Social Choice: From Learning Theory to AI Alignment

Tianyi Qiu

机构 * Peking University(北京大学) UC Berkeley(加州大学伯克利分校) Center for Human-Compatible AI(人类兼容人工智能中心)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments Journal of Artificial Intelligence Research, in press. Best Paper at NeurIPS 2024 Pluralistic Alignment Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 其他LLM 12 篇

2510.27087 2025-11-03 cs.CL cs.CY 89%

Characterizing Selective Refusal Bias in Large Language Models

Adel Khorramrouz, Sharon Levy

机构 * Rutgers University(罗格斯大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

Comments 21 pages, 12 figures, 14 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27190 2025-11-03 cs.CR cs.AI 88%

Unvalidated Trust: Cross-Stage Vulnerabilities in Large Language Model Architectures

Dominik Schwarz

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI;LLM(comments)

Comments 178 pages, mechanism-centered taxonomy of 41 LLM risk patterns, extensive appendix with experiment prompts and consolidation tables. Full traces available to reviewers and affected providers

详情

展开后加载摘要…

URL PDF HTML 收藏