arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-10-02 至 2025-10-02 共收录 192 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 13 篇

2510.00261 2025-10-02 cs.CL cs.AI cs.MM 86%

Retrieval-Augmented Generation for Electrocardiogram-Language Models

Xiaoyu Song, William Han, Tony Chen, Chaojing Duan, Michael A. Rosenberg, Emerson Liu, Ding Zhao

机构 * Carnegie Mellon University(卡内基梅隆大学) Columbia University(哥伦比亚大学) Allegheny Health Network(阿勒格尼医疗网络) University of Colorado(科罗拉多大学)

专题命中 领域大模型 :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL、cs.AI

Comments 5 pages, 2 figures; Submitted to ICASSP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00487 2025-10-02 cs.LG cs.AI 81%

Black-Box Time-Series Domain Adaptation via Cross-Prompt Foundation Models

M. T. Furqon, Mahardhika Pratama, Igor Skrjanc, Lin Liu, Habibullah Habibullah, Kutluyil Dogancay

机构 * STEM, University of South Australia(南澳大利亚大学STEM学院) Faculty of Electrical and Computer Engineering, University of Ljubljana(卢布尔雅那大学电气与计算机工程学院)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00890 2025-10-02 cs.CL cs.AI 79%

Span-level Detection of AI-generated Scientific Text via Contrastive Learning and Structural Calibration

Zhen Yin, Shenghua Wang

机构 * Beijing Renhe Information Technology Co., Ltd.(北京润和信息技术有限公司) Key Laboratory of Digital Publishing and Total Process Management of Scientific and Technical Journals(科技期刊数字出版与全流程管理重点实验室)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16495 2025-10-02 cs.LG cs.AI 79%

Post Hoc Regression Refinement via Pairwise Rankings

Kevin Tirta Wijaya, Michael Sun, Minghao Guo, Hans-Peter Seidel, Wojciech Matusik, Vahid Babaei

机构 * MPI-INF(马克斯·普朗克研究所信息部门) MIT(麻省理工学院)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments NeurIPS 2025 camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.12950 2025-10-02 cs.CL 77%

GuRE:Generative Query REwriter for Legal Passage Retrieval

Daehee Kim, Deokhyung Kang, Jonghwi Kim, Sangwon Ryu, Gary Geunbae Lee

机构 * Graduate School of Artificial Intelligence, POSTECH, Republic of Korea(人工智能研究生院,POSTECH,大韩民国) AI Future Lab, KT, Republic of Korea(AI未来实验室,KT,大韩民国) Department of Computer Science and Engineering, POSTECH, Republic of Korea(计算机科学与工程系,POSTECH,大韩民国)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments NLLP Workshop at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14738 2025-10-02 cs.AI 74%

R&D-Agent: An LLM-Agent Framework Towards Autonomous Data Science

Xu Yang, Xiao Yang, Shikai Fang, Yifei Zhang, Jian Wang, Bowen Xian, Qizheng Li, Jingyuan Li, Minrui Xu, Yuante Li, Haoran Pan, Yuge Zhang, Weiqing Liu, Yelong Shen, Weizhu Chen, Jiang Bian

专题命中 领域大模型 :LLM(title);分类 cs.AI

Comments 33 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00428 2025-10-02 cs.LG cs.AI 73%

Automated Structured Radiology Report Generation with Rich Clinical Context

Seongjae Kang, Dong Bok Lee, Juho Jung, Dongseop Kim, Won Hwa Kim, Sunghoon Joo

机构 * VUNO Inc.(VUNO公司) KAIST(韩国科学技术院) POSTECH(POSTECH大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments 34 pages, 30 figures, preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00129 2025-10-02 cs.LG cond-mat.mtrl-sci cs.AI physics.comp-ph 73%

BigBang-Proton Technical Report: Next-Word-Prediction is Scientific Multitask Learner

Hengkui Wu, Liujiang Liu, Jihua He, Qihao Wang, Keke Zhao, Shuyang Hu, Renle Fu, Dahao Liang, Lingyu Zeng, Bruce Liu, Yuan Liu, Jin Zhan, Jiaqiang Niu, Xinglong Jia, Yaqin Hu, Wenjun Ji, Panpan Chi, Ken Chen, Hengyuan Wu, Yingsi Xin, Yongfeng Zhu, Yuexin Wang, Manqi Ruan, Ningtao Bian, Xiaohua Wu, Weipeng Xu

机构 * SuperSymmetry Technologies(超对称技术)

专题命中 领域大模型 :language model(abstract);pretraining(abstract);分类 cs.AI、cs.LG

Comments 93 pages, 39 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05978 2025-10-02 eess.IV cs.CL cs.CV cs.LG 73%

Imagining Alternatives: Towards High-Resolution 3D Counterfactual Medical Image Generation via Language Guidance

Mohamed Mohamed, Brennan Nichyporuk, Douglas L. Arnold, Tal Arbel

机构 * McGill University(麦吉尔大学) Mila – Quebec AI Institute(魁北克人工智能研究所)

专题命中 领域大模型 :language model(abstract);foundation model(abstract);分类 cs.CL、cs.LG

Comments Accepted to the 2025 MICCAI ELAMI Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00793 2025-10-02 cs.AI cs.CY 70%

AI in data science education: experiences from the classroom

J. A. Hageman, C. F. W. Peeters

机构 * Mathematical and Statistical Methods Group (Biometris)(数学与统计方法组(生物计量学)) Wageningen University(瓦根ingen大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

Comments 6 pages, 0 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.13497 2025-10-02 cs.RO cs.AI 70%

Learning Hierarchical Domain Models Through Environment-Grounded Interaction

Claudius Kienle, Benjamin Alt, Oleg Arenz, Jan Peters

机构 * Intelligent Autonomous Systems Group(智能自主系统组) TU Darmstadt(图宾根大学) AICOR Insitute for Artificial Intelligence(人工智能研究所) University of Bremen(不莱梅大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00585 2025-10-02 eess.IV cs.AI cs.CV 57%

U-DFA: A Unified DINOv2-Unet with Dual Fusion Attention for Multi-Dataset Medical Segmentation

Zulkaif Sajjad, Furqan Shaukat, Junaid Mir

机构 * Department of Electrical and Electronics Engineering(电气与电子工程系) University of Engineering and Technology Taxila(工程与技术大学)

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.22678 2025-10-02 cs.CL 57%

Self-Evolving Multi-Agent Simulations for Realistic Clinical Interactions

Mohammad Almansoori, Komal Kumar, Hisham Cholakkal

机构 * Mohamed bin Zayed University of Artificial Intelligence(Mohamed bin Zayed人工智能大学)

专题命中 领域大模型 :LLM(abstract);分类 cs.CL

Comments 14 page, 4 figures, 61 references, presented in MICCAI (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 12 篇

2507.19419 2025-10-02 cs.CL 85%

TokenSmith: Streamlining Data Editing, Search, and Inspection for Large-Scale Language Model Training and Interpretability

Mohammad Aflah Khan, Ameya Godbole, Johnny Tian-Zheng Wei, Ryan Wang, James Flemings, Krishna P. Gummadi, Willie Neiswanger, Robin Jia

机构 * Max Planck Institute for Software Systems(马克斯·普朗克软件系统研究所) University of Southern California(南加州大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);pretraining(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17349 2025-10-02 cs.CV 82%

Beyond Semantics: Rediscovering Spatial Awareness in Vision-Language Models

Jianing Qi, Jiawei Liu, Hao Tang, Zhigang Zhu

机构 * CUNY Graduate Center(纽约大学研究生中心) Borough of Manhattan Community College(曼哈顿社区学院) The City College of New York(纽约城市学院)

专题命中 知识编辑与模型理解 :language model(title,abstract);LLM(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.12838 2025-10-02 cs.CL 79%

Are Knowledge and Reference in Multilingual Language Models Cross-Lingually Consistent?

Xi Ai, Mahardika Krisna Ihsani, Min-Yen Kan

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments EMNLP'25 Findings Camera Ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11679 2025-10-02 cs.CL cs.LG 79%

Ambiguity in LLMs is a concept missing problem

Zhibo Hu, Chen Wang, Yanfeng Shu, Hye-Young Paik, Liming Zhu

机构 * The University of New South Wales(新南威尔士大学) CSIRO Data61(澳大利亚联邦科学与工业研究组织数据61)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments 17 pages, 11 figures, title updated

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17324 2025-10-02 cs.CL cs.AI 77%

CultranAI at PalmX 2025: Data Augmentation for Cultural Knowledge Representation

Hunzalah Hassan Bhatti, Youssef Ahmed, Md Arid Hasan, Firoj Alam

专题命中 知识编辑与模型理解 :large language model(abstract,comments);language model(abstract,comments);分类 cs.CL、cs.AI;foundation model(comments)

Comments LLMs, Native, Arabic LLMs, Augmentation, Multilingual, Language Diversity, Contextual Understanding, Minority Languages, Culturally Informed, Foundation Models, Large Language Models

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00296 2025-10-02 cs.LG 77%

Beyond Token Probes: Hallucination Detection via Activation Tensors with ACT-ViT

Guy Bar-Shalom, Fabrizio Frasca, Yaniv Galron, Yftah Ziser, Haggai Maron

机构 * Technion(技术离子大学) University of Groningen(格罗宁根大学) Nvidia Research(英伟达研究)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

Comments Published in NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11895 2025-10-02 cs.CL cs.AI cs.LG 75%

Resolving UnderEdit & OverEdit with Iterative & Neighbor-Assisted Model Editing

Bhiman Kumar Baghel, Emma Jordan, Zheyuan Ryan Shi, Xiang Lorraine Li

机构 * Department of Computer Science, University of Pittsburgh, PA, USA(计算机科学系,匹兹堡大学,宾夕法尼亚州,美国)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted at EMNLP 2025 as Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00625 2025-10-02 cs.AI 70%

Is Model Editing Built on Sand? Revealing Its Illusory Success and Fragile Foundation

Wei Liu, Haomei Xu, Bingqing Liu, Zhiying Deng, Haozhao Wang, Jun Wang, Ruixuan Li, Yee Whye Teh, Wee Sun Lee

机构 * National University of Singapore(新加坡国立大学) Huazhong University of Science and Technology(华中科技大学) Central China Normal University(中国地质大学) iWudao Tech(iWudao科技) Oxford(牛津大学) Google Deepmind(谷歌DeepMind)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

Comments This is a work in progress. Comments and suggestions are welcome

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01004 2025-10-02 cs.CV cs.AI cs.LG 62%

TextCAM: Explaining Class Activation Map with Text

Qiming Zhao, Xingjian Li, Xiaoyu Cao, Xiaolong Wu, Min Xu

机构 * Carnegie Mellon University(卡内基梅隆大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.19411 2025-10-02 cs.IR cs.CL cs.LG 62%

PaECTER: Patent-level Representation Learning using Citation-informed Transformers

Mainak Ghosh, Michael E. Rose, Sebastian Erhardt, Erik Buunk, Dietmar Harhoff

机构 * Max Planck Institute for Innovation and Competition(马克斯·普朗克创新与竞争研究所)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.LG

Comments 8 pages, 3 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00721 2025-10-02 physics.chem-ph 50%

Flexible Uncertainty Calibration for Machine-Learned Interatomic Potentials

Cheuk Hin Ho, Christoph Ortner, Yangshuai Wang

专题命中 知识编辑与模型理解 :foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00683 2025-10-02 cs.CV 50%

ProtoMask: Segmentation-Guided Prototype Learning

Steffen Meinert, Philipp Schlinge, Nils Strodthoff, Martin Atzmueller

机构 * AI4Health Division, Faculty VI Carl von Ossietzky Universität Oldenburg(AI4Health部门,第六学院卡尔·冯·奥西特齐克大学) Semantic Information Systems Group Osnabrück University(语义信息系统小组奥尔纳布里克大学)

专题命中 知识编辑与模型理解 :foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他LLM 17 篇

2510.00047 2025-10-02 cs.CV cs.AI 83%

Explanation-Driven Counterfactual Testing for Faithfulness in Vision-Language Model Explanations

Sihao Ding, Santosh Vasa, Aditi Ramadwar

机构 * Mercedes-Benz Research & Development North America(梅赛德斯-奔驰研究与开发北美)

专题命中 其他LLM :language model(title,abstract);LLM(abstract);分类 cs.AI

Comments NeurIPS 2025 workshop on Regulatable ML

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26181 2025-10-02 cs.CL 83%

Explaining novel senses using definition generation with open language models

Mariia Fedorova, Andrey Kutuzov, Francesco Periti, Yves Scherrer

机构 * University of Oslo(奥斯陆大学) KU Leuven - Flanders Make(库尔勒文大学-佛兰德斯制造)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL

Comments EMNLP 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00276 2025-10-02 cs.CL cs.LG 79%

SafePassage: High-Fidelity Information Extraction with Black Box LLMs

Joe Barrow, Raj Patel, Misha Kharkovski, Ben Davies, Ryan Schmitt

机构 * Pattern Data

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.20485 2025-10-02 cs.CR cs.CL cs.LG 79%

Phantom: General Backdoor Attacks on Retrieval Augmented Language Generation

Harsh Chaudhari, Giorgio Severi, John Abascal, Anshuman Suri, Matthew Jagielski, Christopher A. Choquette-Choo, Milad Nasr, Cristina Nita-Rotaru, Alina Oprea

机构 * Northeastern University(东北大学) OpenAI

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01088 2025-10-02 cs.AI 77%

Safety Instincts: LLMs Learn to Trust Their Internal Compass for Self-Defense

Guobin Shen, Dongcheng Zhao, Haibo Tong, Jindong Li, Feifei Zhao, Yi Zeng

机构 * Beijing Institute of AI Safety and Governance(北京人工智能安全与治理研究院) Beijing Key Laboratory of Safe AI and Superalignment(北京安全人工智能与超对齐重点实验室) BrainCog Lab, Institute of Automation, Chinese Academy of Sciences(脑认知实验室,中国科学院自动化研究所)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏