arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-09-25 至 2025-09-25 共收录 179 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 19 篇

2509.19090 2025-09-25 cs.CV cs.AI cs.CL 81%

Citrus-V: Advancing Medical Foundation Models with Unified Medical Image Grounding for Clinical Reasoning

Guoxin Wang, Jun Zhao, Xinyi Liu, Yanbo Liu, Xuyang Cao, Chao Li, Zhuoyun Liu, Qintian Sun, Fangru Zhou, Haoqiang Xing, Zhenhong Yang

机构 * JDH Algo(京东健康算法)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19925 2025-09-25 cs.AI 77%

CON-QA: Privacy-Preserving QA using cloud LLMs in Contract Domain

Ajeet Kumar Singh, Rajsabi Surya, Anurag Tripathi, Santanu Choudhury, Sudhir Bisane

机构 * Info Origin INC Indian Institute of Technology New Delhi (IITD)(印度理工学院新德里学院(IITD))

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19325 2025-09-25 cs.CL 77%

How Much of Your Data Can Suck? Thresholds for Domain Performance and Emergent Misalignment in LLMs

Jian Ouyang, Arman T, Ge Jin

机构 * Invisible Technologies(隐形技术)

专题命中 领域大模型 :large language model(abstract);language model(abstract);SFT(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.01841 2025-09-25 eess.AS cs.AI cs.CL cs.IR cs.SD 73%

A GEN AI Framework for Medical Note Generation

Hui Yi Leong, Yi Fan Gao, Shuai Ji, Bora Kalaycioglu, Uktu Pamuksuz

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 8 Figures, 7 page, IEEE standard research paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08742 2025-09-25 cs.CL cs.AI 73%

SciRerankBench: Benchmarking Rerankers Towards Scientific Retrieval-Augmented Generated LLMs

Haotian Chen, Qingqing Long, Meng Xiao, Xiao Luo, Wei Ju, Chengrui Wang, Xuezhi Wang, Yuanchun Zhou, Hengshu Zhu

机构 * Computer Network Information Center, Chinese Academy of Sciences(中国科学院计算机网络信息中心) University of California, Los Angeles(加州大学洛杉矶分校) Peking University(北京大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.20575 2025-09-25 cs.LG 57%

Exploring Graph-Transformer Out-of-Distribution Generalization Abilities

Itay Niv, Neta Rabin

机构 * Faculty of Engineering(工程学院) Tel-Aviv University(特拉维夫大学)

专题命中 领域大模型 :post-training(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23673 2025-09-25 cs.CV cs.LG 57%

SAMSA: Segment Anything Model Enhanced with Spectral Angles for Hyperspectral Interactive Medical Image Segmentation

Alfie Roddan, Tobias Czempiel, Chi Xu, Daniel S. Elson, Stamatia Giannarou

机构 * The Hamlyn Centre for Robotic Surgery, Imperial College London, UK(哈姆林机器人手术中心,帝国理工学院伦敦分校)

专题命中 领域大模型 :foundation model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20279 2025-09-25 cs.CV q-bio.QM 50%

A co-evolving agentic AI system for medical imaging analysis

Songhao Li, Jonathan Xu, Tiancheng Bao, Yuxuan Liu, Yuchen Liu, Yihang Liu, Lilin Wang, Wenhui Lei, Sheng Wang, Yinuo Xu, Yan Cui, Jialu Yao, Shunsuke Koga, Zhi Huang

机构 * Department of Pathology and Laboratory Medicine, University of Pennsylvania(病理学与实验室医学系,宾夕法尼亚大学) Department of Electrical and System Engineering, University of Pennsylvania(电气与系统工程系,宾夕法尼亚大学) The Wharton School, University of Pennsylvania(沃顿商学院,宾夕法尼亚大学) Department of Bioengineering, University of Pennsylvania(生物工程系,宾夕法尼亚大学) Department of Computer and Information Science, University of Pennsylvania(计算机与信息科学系,宾夕法尼亚大学) Department of Biostatistics, Epidemiology & Informatics, University of Pennsylvania(生物统计学、流行病学与信息学系,宾夕法尼亚大学)

专题命中 领域大模型 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20119 2025-09-25 cs.CV 50%

A Simple Data Augmentation Strategy for Text-in-Image Scientific VQA

Belal Shoer, Yova Kementchedjhieva

机构 * MBZUAI(穆罕默德·本·拉希德智能研究院)

专题命中 领域大模型 :language model(abstract)

Comments Accepted at WiNLP, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 11 篇

2509.19375 2025-09-25 cs.LG cs.AI stat.ML 88%

Uncertainty Quantification of Large Language Models using Approximate Bayesian Computation

Mridul Sharma, Adeetya Patel, Zaneta D' Souza, Samira Abbasgholizadeh Rahimi, Siva Reddy, Sreenath Madathil

机构 * Faculty of Dental Medicine and Oral Health Sciences, McGill University(牙医学院与口腔健康科学学院,麦吉尔大学) McGill University(麦吉尔大学) Mila–Quebec Artificial Intelligence Institute(魁北克人工智能研究所) School of Computer Science, McGill University(计算机科学学院,麦吉尔大学)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19563 2025-09-25 cs.CL cs.LG 81%

Uncertainty in Semantic Language Modeling with PIXELS

Stefania Radu, Marco Zullich, Matias Valdenegro-Toro

机构 * Department of Artificial Intelligence, Bernoulli Institute, University of Groningen(人工智能系、伯努利研究所、 Groningen大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.LG

Comments 9 pages, 6 figures, UncertaiNLP 2025 Workshop @ EMNLP Camera Ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.18485 2025-09-25 q-bio.NC cs.CV 78%

Deciphering Functions of Neurons in Vision-Language Models

Jiaqi Xu, Cuiling Lan, Yan Lu

机构 * University of Science and Technology of China(中国科学技术大学) Microsoft Research Asia(微软亚洲研究院)

专题命中 知识编辑与模型理解 :language model(title,abstract)

Comments Accepted by the 31st ACM International Conference on Multimedia (ACM MM 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19839 2025-09-25 cs.AI 77%

LatentGuard: Controllable Latent Steering for Robust Refusal of Attacks and Reliable Response Generation

Huizhen Shu, Xuying Li, Zhuo Li

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments 9-page NeurIPS 2025 preprint including 3 figures and 1 table, with additional appendix material. Prepared using the NeurIPS 2025 preprint template and compiled with pdfLaTeX. All references are included via the provided .bbl file. Figures are in PDF format. No external supplementary files. All necessary style files and images are included

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19745 2025-09-25 cs.CL cs.SD 77%

PART: Progressive Alignment Representation Training for Multilingual Speech-To-Text with LLMs

Pei Zhang, Andong Chen, Xi Chen, Baosong Yang, Derek F. Wong, Fei Huang

机构 * Tongyi Lab, Alibaba Group(通义实验室,阿里巴巴集团) The Chinese University of Hong Kong(香港中文大学) NLP 2 CT Lab, University of Macau(自然语言处理2CT实验室,澳门大学)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.03090 2025-09-25 cs.CL cs.LG 73%

UNComp: Can Matrix Entropy Uncover Sparsity? -- A Compressor Design from an Uncertainty-Aware Perspective

Jing Xiong, Jianghan Shen, Fanghua Ye, Chaofan Tao, Zhongwei Wan, Jianqiao Lu, Xun Wu, Chuanyang Zheng, Zhijiang Guo, Min Yang, Lingpeng Kong, Ngai Wong

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments Accepted at EMNLP 2025 (Main Conference)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20168 2025-09-25 cs.CL 70%

Probing Gender Bias in Multilingual LLMs: A Case Study of Stereotypes in Persian

Ghazal Kalhor, Behnam Bahrak

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted and forthcoming at the Widening Natural Language Processing Workshop (WiNLP 2025) at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20088 2025-09-25 cs.CL cs.AI 62%

Causal Understanding by LLMs: The Role of Uncertainty

Oscar Lithgow-Serrano, Vani Kanjirangat, Alessandro Antonucci

机构 * SUPSI, IDSIA(瑞士SUPSI和IDSIA)

专题命中 知识编辑与模型理解 :pretraining(abstract);分类 cs.CL、cs.AI

Comments Accepted in second UncertaiNLP workshop at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20065 2025-09-25 cs.CL 57%

From Input Perception to Predictive Insight: Modeling Model Blind Spots Before They Become Errors

Maggie Mi, Aline Villavicencio, Nafise Sadat Moosavi

机构 * University of Sheffield(谢菲尔德大学) University of Exeter(埃克塞特大学) The Alan Turing Institute(艾伦·图灵研究所) UFRN, Brazil(巴西UFRN)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.13390 2025-09-25 cs.CL 57%

Aligned Probing: Relating Toxic Behavior and Model Internals

Andreas Waldis, Vagrant Gautam, Anne Lauscher, Dietrich Klakow, Iryna Gurevych

机构 * Ubiquitous Knowledge Processing Lab (UKP Lab)(通用知识处理实验室) Technical University of Darmstadt(德累斯顿技术大学) Information Systems Research Lab(信息系统研究实验室) Lucerne University of Applied Sciences and Arts(卢塞恩应用科学与艺术大学) Spoken Language Systems(语音语言系统) Saarland University(萨尔兰大学) Data Science Group(数据科学组) University of Hamburg(汉堡大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19562 2025-09-25 cs.CV 50%

CURE: Centroid-guided Unsupervised Representation Erasure for Facial Recognition Systems

Fnu Shivam, Nima Najafzadeh, Yenumula Reddy, Prashnna Gyawali

机构 * West Virginia University(西弗吉尼亚大学)

专题命中 知识编辑与模型理解 :prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他LLM 9 篇

2509.19727 2025-09-25 cs.CL 88%

Personality Vector: Modulating Personality of Large Language Models by Model Merging

Seungjong Sun, Seo Yeon Baek, Jang Hyun Kim

机构 * Department of Human-Artificial Intelligence Interaction, Sungkyunkwan University(人类-人工智能交互系,成均馆大学) Department of Immersive Media Engineering, Sungkyunkwan University(沉浸式媒体工程系,成均馆大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07513 2025-09-25 cs.CL cs.AI 86%

Language Models Fail to Introspect About Their Knowledge of Language

Siyuan Song, Jennifer Hu, Kyle Mahowald

机构 * Department of Linguistics The University of Texas at Austin(语言学系 德克萨斯大学奥斯汀分校) Department of Cognitive Science Johns Hopkins University(认知科学系 约翰霍普金斯大学)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);prompting(abstract);分类 cs.CL、cs.AI

Comments 23 pages, 10 figures, COLM 2025 camera ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.12038 2025-09-25 cs.LG 77%

LLMs for Cold-Start Cutting Plane Separator Configuration

Connor Lawless, Yingxi Li, Anders Wikum, Madeleine Udell, Ellen Vitercik

机构 * Management Science and Engineering, Stanford University(管理科学与工程系,斯坦福大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05473 2025-09-25 cs.MM cs.AI cs.SD eess.AS 70%

Embedding Alignment in Code Generation for Audio

Sam Kouteili, Hiren Madhu, George Typaldos, Mark Santolucito

机构 * Yale University(耶鲁大学) Columbia University(哥伦比亚大学)

专题命中 其他LLM :LLM(abstract);prompting(abstract);分类 cs.AI

Comments Accepted to NeurIPS 2025 AI4Music Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23906 2025-09-25 cs.DS cs.CC cs.DC 67%

Segmented Operations using Matrix Multiplications

Aleksandros Sobczyk, Giuseppe Sorrentino, Anastasios Zouzias

专题命中 其他LLM :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19554 2025-09-25 cs.LG cs.AI 62%

Learning Dynamics of Deep Learning -- Force Analysis of Deep Neural Networks

Yi Ren

机构 * The University Of British Columbia(不列颠哥伦比亚大学)

专题命中 其他LLM :LLM(abstract);分类 cs.AI、cs.LG

Comments 175 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.05810 2025-09-25 cs.LG cs.AI physics.chem-ph 62%

A Transformer Model for Predicting Chemical Products from Generic SMARTS Templates with Data Augmentation

Derin Ozer, Sylvain Lamprier, Thomas Cauchy, Nicolas Gutowski, Benoit Da Mota

专题命中 其他LLM :language model(abstract);分类 cs.AI、cs.LG

Comments ICTAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19688 2025-09-25 cs.RO cs.LG cs.SY eess.SY math.OC 57%

Formal Safety Verification and Refinement for Generative Motion Planners via Certified Local Stabilization

Devesh Nath, Haoran Yin, Glen Chou

机构 * Georgia Institute of Technology(佐治亚理工学院)

专题命中 其他LLM :language model(abstract);分类 cs.LG

Comments 10 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13963 2025-09-25 math.RA 50%

Anti-pre-Poisson bialgebras and relative Rota-Baxter operators

Qinxiu Sun, Min Wu

专题命中 其他LLM :prompting(abstract)

Comments 34 pages

详情

展开后加载摘要…

URL PDF HTML 收藏