arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12705 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 12705 篇

2603.12816 2026-03-16 cs.LG cs.AI cs.CV 76%

Residual SODAP: Residual Self-Organizing Domain-Adaptive Prompting with Structural Knowledge Preservation for Continual Learning

残差SODAP:基于结构知识保留的残差自组织领域自适应提示法用于持续学习

Gyutae Oh, Jungwoo Bae, Jitae Shin

机构 * Department of Electrical and Computer Engineering, Sungkyunkwan University, Suwon 16419, Republic of Korea(釜山大学电气与计算机工程系,Suwon 16419,韩国)

专题命中 领域大模型 :prompting(title);分类 cs.AI、cs.LG

AI总结 本文提出残差SODAP,结合提示表示适应与分类器级知识保留,通过α-entmax稀疏提示选择、残差聚合、数据无关蒸馏与伪特征重放等方法,在无任务ID和额外数据存储的三个DIL基准上取得最优AvgACC/AvgF性能。

Comments 29 page, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.09469 2026-02-11 cs.CL cs.AI 76%

NOWJ @BioCreative IX ToxHabits: An Ensemble Deep Learning Approach for Detecting Substance Use and Contextual Information in Clinical Texts

NOWJ @BioCreative IX ToxHabits: 一种用于检测临床文本中物质使用及上下文信息的集成深度学习方法

Huu-Huy-Hoang Tran, Gia-Bao Duong, Quoc-Viet-Anh Tran, Thi-Hai-Yen Vuong, Hoang-Quynh Le

机构 * University of Engineering and Technology(工程大学) University of Engineering(工程大学) Technology, Vietnam National University(技术,越南国家大学)

专题命中 领域大模型 :large language model(abstract,journal_ref);language model(abstract,journal_ref);分类 cs.CL、cs.AI

AI总结 本文提出了一种集成深度学习方法,用于在西班牙临床文本中检测有毒物质使用及上下文信息,实现了较高的F1和精度指标。

Journal ref Proceedings of the BioCreative IX Challenge and Workshop (BC9): Large Language Models for Clinical and Biomedical NLP at the International Joint Conference on Artificial Intelligence (IJCAI 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.15301 2026-01-28 cs.CL cs.AI 76%

Can We Trust LLM Detectors?

我们能相信LLM检测器吗?

Jivnesh Sandhan, Harshit Jaiswal, Fei Cheng, Yugo Murawaki

机构 * Kyoto University(京都大学) IIT Kanpur(印度理工学院坎浦尔)

专题命中 领域大模型 :LLM(title);分类 cs.CL、cs.AI

AI总结 本文提出监督对比学习框架,揭示现有LLM检测器在分布偏移和风格扰动下的脆弱性,指出构建领域无关检测器的挑战。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09715 2026-01-27 cs.CL cs.AI cs.HC cs.IR 76%

Introducing Axlerod: An LLM-based Chatbot for Assisting Independent Insurance Agents

引入Axlerod:一种基于大语言模型的聊天机器人,用于协助独立保险代理

Adam Bradley, John Hastings, Khandaker Mamun Ahmed

机构 * The Beacom College of Computer and Cyber Sciences, Dakota State University(计算机与网络安全学院,达科他州立大学)

专题命中 领域大模型 :LLM(title);分类 cs.CL、cs.AI

AI总结 Axlerod是一种基于大语言模型的聊天机器人,旨在通过自然语言处理、检索增强生成和领域知识整合,提高独立保险代理的工作效率。

Comments 6 pages, 2 figures, 1 table

Journal ref 2025 IEEE Cyber Awareness and Research Symposium (CARS'25)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16444 2026-01-21 cs.AI cs.LG 76%

Domain-Specific Constitutional AI: Enhancing Safety in LLM-Powered Mental Health Chatbots

领域特定的宪法AI:增强LLM驱动的心理健康聊天机器人安全性

Chenhan Lyu, Yutong Song, Pengfei Zhang, Amir M. Rahmani

专题命中 领域大模型 :LLM(title);分类 cs.AI、cs.LG

AI总结 本文提出利用领域特定心理健康原则的宪法AI训练方法,以提升心理健康聊天机器人在安全性和领域适应性方面的表现。

Comments Accepted to 2025 IEEE 21st International Conference on Body Sensor Networks (BSN)

Journal ref 2025 IEEE 21st International Conference on Body Sensor Networks (BSN), pp. 1-4

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.19546 2025-11-26 cs.AI cs.CL cs.HC 76%

CNS-Obsidian: A Neurosurgical Vision-Language Model Built From Scientific Publications

CNS-Obsidian:基于科学出版物构建的神经外科视觉-语言模型

Anton Alyakin, Jaden Stryker, Daniel Alexander Alber, Jin Vivian Lee, Karl L. Sangwon, Brandon Duderstadt, Akshay Save, David Kurland, Spencer Frome, Shrutika Singh, Jeff Zhang, Eunice Yang, Ki Yun Park, Cordelia Orillac, Aly A. Valliani, Sean Neifert, Albert Liu, Aneek Patel, Christopher Livia, Darryl Lau, Ilya Laufer, Peter A. Rozman, Eveline Teresa Hidalgo, Howard Riina, Rui Feng, Todd Hollon, Yindalon Aphinyanaphongs, John G. Golfinos, Laura Snyder, Eric Leuthardt, Douglas Kondziolka, Eric Karl Oermann

专题命中 领域大模型 :language model(title);分类 cs.CL、cs.AI

AI总结 CNS-Obsidian是一款基于科学文献训练的神经外科视觉-语言模型,通过对比实验展示了其在诊断准确性与用户评分上的表现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19224 2025-09-26 cs.CL cs.AI 76%

Systematic Comparative Analysis of Large Pretrained Language Models on Contextualized Medication Event Extraction

Tariq Abdul-Quddoos, Xishuang Dong, Lijun Qian

机构 * Center of Excellence in Research and Education for Big Military Data Intelligence (CREDIT Center)(大数据智能卓越研究中心(CREDIT中心)) Department of Electrical and Computer Engineering(电气与计算机工程系) Prairie View A&M University(普拉提维大学) Texas A&M University System(德克萨斯A&M大学系统)

专题命中 领域大模型 :language model(title);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04982 2025-09-08 cs.CL cs.IR cs.LG 76%

Optimizing Small Transformer-Based Language Models for Multi-Label Sentiment Analysis in Short Texts

Julius Neumann, Robert Lange, Yuni Susanti, Michael Färber

机构 * ScaDS.AI, TU Dresden, Germany(ScaDS.AI,图腾德斯大学,德国) FIZ Karlsruhe, Germany(卡尔斯鲁厄研究所,德国)

专题命中 领域大模型 :language model(title);分类 cs.CL、cs.LG

Comments Accepted at LDD@ECAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.06124 2025-08-07 cs.LG cs.AI 76%

Foundation Model of Electronic Medical Records for Adaptive Risk Estimation

Pawel Renc, Michal K. Grzeszczyk, Nassim Oufattole, Deirdre Goode, Yugang Jia, Szymon Bieganski, Matthew B. A. McDermott, Jaroslaw Was, Anthony E. Samir, Jonathan W. Cunningham, David W. Bates, Arkadiusz Sitek

机构 * Massachusetts General Hospital(麻省总医院) Harvard Medical School(哈佛医学院) AGH University of Krakow(克拉科夫AGH大学) Massachusetts Institute of Technology(麻省理工学院) Newton Wellesley Hospital(牛顿韦尔士医院) Medical University of Lodz(罗兹医学院) Harvard Chan School of Public Health(哈佛医学院公共卫生学院)

专题命中 领域大模型 :foundation model(title);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02999 2025-08-06 cs.AI cs.CL 76%

AGENTiGraph: A Multi-Agent Knowledge Graph Framework for Interactive, Domain-Specific LLM Chatbots

Xinjie Zhao, Moritz Blum, Fan Gao, Yingjian Chen, Boming Yang, Luis Marquez-Carpintero, Mónica Pina-Navarro, Yanran Fu, So Morikawa, Yusuke Iwasawa, Yutaka Matsuo, Chanjun Park, Irene Li

机构 * The University of Tokyo(东京大学) University of Bielefeld(比勒菲尔德大学) University of Alicante(阿利坎特大学) Soongsil University(顺世大学)

专题命中 领域大模型 :LLM(title);分类 cs.CL、cs.AI

Comments CIKM 2025, Demo Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18263 2025-07-25 cs.CL cs.AI 76%

Locate-and-Focus: Enhancing Terminology Translation in Speech Language Models

Suhang Wu, Jialong Tang, Chengyi Yang, Pei Zhang, Baosong Yang, Junhui Li, Junfeng Yao, Min Zhang, Jinsong Su

机构 * Department of Digital Media Technology, Xiamen University(厦门大学数字媒体技术系) Tongyi Lab(通义实验室) Soochow University(苏州大学) Key Laboratory of Digital Protection and Intelligent Processing of Intangible Cultural Heritage of Fujian and Taiwan (Xiamen University), Ministry of Culture and Tourism, China(福建省和台湾非物质文化遗产数字化保护与智能处理重点实验室(厦门大学),中华人民共和国文化和旅游部,中国)

专题命中 领域大模型 :language model(title);分类 cs.CL、cs.AI

Comments Accepted at ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.18986 2025-04-25 cs.CL cs.AI cs.IR 76%

Lab-AI: Using Retrieval Augmentation to Enhance Language Models for Personalized Lab Test Interpretation in Clinical Medicine

Xiaoyu Wang, Haoyong Ouyang, Balu Bhasuran, Xiao Luo, Karim Hanna, Mia Liza A. Lustria, Carl Yang, Zhe He

机构 * Department of Statistics(统计学系) Florida State University(佛罗里达州立大学) School of Information(信息学院) Oklahoma State University(俄克拉荷马州立大学) Morsani College of Medicine(医学学院) University of South Florida(佛罗里达大学) College of Social Work(社会工作学院) Department of Computer Science(计算机科学系) Emory University(埃默里大学)

专题命中 领域大模型 :language model(title);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17382 2025-03-25 cs.LG cs.CL 76%

State Fourier Diffusion Language Model (SFDLM): A Scalable, Novel Iterative Approach to Language Modeling

Andrew Kiruluta, Andreas Lemos

机构 * University of California Berkeley(加州大学伯克利分校)

专题命中 领域大模型 :language model(title);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.06084 2025-02-11 cs.LG cs.AI cs.NE 76%

Physics-Guided Foundation Model for Scientific Discovery: An Application to Aquatic Science

Runlong Yu, Chonghao Qiu, Robert Ladwig, Paul Hanson, Yiqun Xie, Xiaowei Jia

专题命中 领域大模型 :foundation model(title);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.08243 2025-01-15 cs.SE cs.AI cs.LG 76%

Engineering LLM Powered Multi-agent Framework for Autonomous CloudOps

Kannan Parthasarathy, Karthik Vaidhyanathan, Rudra Dhar, Venkat Krishnamachari, Basil Muhammed, Adyansh Kakran, Sreemaee Akshathala, Shrikara Arun, Sumant Dubey, Mohan Veerubhotla, Amey Karan

专题命中 领域大模型 :LLM(title);分类 cs.AI、cs.LG

Comments The paper has been accepted as full paper to CAIN 2025 (https://conf.researchr.org/home/cain-2025), co-located with ICSE 2025 (https://conf.researchr.org/home/icse-2025). The paper was submitted to CAIN for review on 9 November 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.00767 2024-12-03 cs.CV cs.CL cs.LG 76%

Prompt as Free Lunch: Enhancing Diversity in Source-Free Cross-domain Few-shot Learning through Semantic-Guided Prompting

Linhai Zhuo, Zheng Wang, Yuqian Fu, Tianwen Qian

专题命中 领域大模型 :prompting(title);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12861 2024-11-05 cs.CL cs.AI cs.HC 76%

CiteME: Can Language Models Accurately Cite Scientific Claims?

Ori Press, Andreas Hochlehnert, Ameya Prabhu, Vishaal Udandarao, Ofir Press, Matthias Bethge

专题命中 领域大模型 :language model(title);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.14180 2024-11-04 cs.CL cs.LG stat.AP stat.ML 76%

Dynamic Topic Language Model on Heterogeneous Children's Mental Health Clinical Notes

Hanwen Ye, Tatiana Moreno, Adrianne Alpern, Louis Ehwerhemuepha, Annie Qu

专题命中 领域大模型 :language model(title);分类 cs.CL、cs.LG

Journal ref Ann. Appl. Stat. 18(4): 3165-3184 (December 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.12192 2024-11-01 cs.RO cs.AI cs.CV cs.LG 76%

DynaMo: In-Domain Dynamics Pretraining for Visuo-Motor Control

Zichen Jeff Cui, Hengkai Pan, Aadhithya Iyer, Siddhant Haldar, Lerrel Pinto

专题命中 领域大模型 :pretraining(title);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.03640 2024-10-15 cs.CL cs.AI 76%

Apollo: A Lightweight Multilingual Medical LLM towards Democratizing Medical AI to 6B People

Xidong Wang, Nuo Chen, Junyin Chen, Yidong Wang, Guorui Zhen, Chunxian Zhang, Xiangbo Wu, Yan Hu, Anningzhe Gao, Xiang Wan, Haizhou Li, Benyou Wang

专题命中 领域大模型 :LLM(title);分类 cs.CL、cs.AI

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.14302 2024-10-03 cs.CL cs.AI 76%

Reliable and diverse evaluation of LLM medical knowledge mastery

Yuxuan Zhou, Xien Liu, Chen Ning, Xiao Zhang, Ji Wu

专题命中 领域大模型 :LLM(title);分类 cs.CL、cs.AI

Comments 20 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.18998 2024-10-01 cs.CL cs.AI 76%

Controlled LLM-based Reasoning for Clinical Trial Retrieval

Mael Jullien, Alex Bogatu, Harriet Unsworth, Andre Freitas

专题命中 领域大模型 :LLM(title);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.09395 2024-07-03 q-bio.NC cs.AI cs.CL 76%

Matching domain experts by training from scratch on domain knowledge

Xiaoliang Luo, Guangzhi Sun, Bradley C. Love

专题命中 领域大模型 :large language model(abstract,comments);language model(abstract,comments);分类 cs.CL、cs.AI

Comments ICML 2024 (Large Language Models and Cognition)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.18421 2024-03-28 cs.CL cs.AI 76%

BioMedLM: A 2.7B Parameter Language Model Trained On Biomedical Text

Elliot Bolton, Abhinav Venigalla, Michihiro Yasunaga, David Hall, Betty Xiong, Tony Lee, Roxana Daneshjou, Jonathan Frankle, Percy Liang, Michael Carbin, Christopher D. Manning

专题命中 领域大模型 :language model(title);分类 cs.CL、cs.AI

Comments 23 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.07340 2023-08-15 cs.CL cs.AI 76%

MathBERT: A Pre-trained Language Model for General NLP Tasks in Mathematics Education

Jia Tracy Shen, Michiharu Yamashita, Ethan Prihar, Neil Heffernan, Xintao Wu, Ben Graff, Dongwon Lee

专题命中 领域大模型 :language model(title);分类 cs.CL、cs.AI

Comments Accepted by NeurIPS 2021 MATHAI4ED Workshop (Best Paper)

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.12675 2023-06-01 cs.LG cs.CL cs.CR 76%

Decepticons: Corrupted Transformers Breach Privacy in Federated Learning for Language Models

Liam Fowl, Jonas Geiping, Steven Reich, Yuxin Wen, Wojtek Czaja, Micah Goldblum, Tom Goldstein

专题命中 领域大模型 :language model(title);分类 cs.CL、cs.LG

Comments First two authors contributed equally. Order chosen by coin flip. Published at ICLR 2023. Implementation available at github.com/JonasGeiping/breaching

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.03488 2023-05-04 cs.CL cs.AI 76%

APAM: Adaptive Pre-training and Adaptive Meta Learning in Language Model for Noisy Labels and Long-tailed Learning

Sunyi Chi, Bo Dong, Yiming Xu, Zhenyu Shi, Zheng Du

专题命中 领域大模型 :language model(title);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.01217 2022-12-05 cs.CL cs.AI 76%

Using Large Pre-Trained Language Model to Assist FDA in Premarket Medical Device

Zongzhe Xu

专题命中 领域大模型 :language model(title);分类 cs.CL、cs.AI

Comments IEEE Southeast Conference 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.14812 2022-09-30 cs.AI cs.CL 76%

Named Entity Recognition in Industrial Tables using Tabular Language Models

Aneta Koleva, Martin Ringsquandl, Mark Buckley, Rakebul Hasan, Volker Tresp

专题命中 领域大模型 :language model(title);分类 cs.CL、cs.AI

Comments EMNLP 2022 Industry Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.03951 2022-04-11 cs.CL cs.AI 76%

RuBioRoBERTa: a pre-trained biomedical language model for Russian language biomedical text mining

Alexander Yalunin, Alexander Nesterov, Dmitriy Umerenkov

专题命中 领域大模型 :language model(title);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏