arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12581 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 12581 篇

2604.18786 2026-04-22 cs.CL cs.AI 91%

Experiments or Outcomes? Probing Scientific Feasibility in Large Language Models

实验还是结果?在大型语言模型中探测科学可行性

Seyedali Mohammadi, Manas Gaur, Francis Ferraro

机构 * University of Maryland, Baltimore County(马里兰大学巴尔的摩分校)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn);分类 cs.CL、cs.AI

AI总结 本文探讨了在大型语言模型中科学可行性的评估,发现提供结果证据比实验描述更可靠,结果能提升准确性,而实验文本可能因上下文不完整而退化。

Comments Accepted at ACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21248 2026-04-21 cs.CL cs.AI cs.CE 91%

ResearchBench: Benchmarking LLMs in Scientific Discovery via Inspiration-Based Task Decomposition

ResearchBench: 通过基于灵感的任务分解对LLM在科学发现中的基准测试

Yujie Liu, Zonglin Yang, Tong Xie, Jinjie Ni, Ben Gao, Yuqiang Li, Shixiang Tang, Wanli Ouyang, Erik Cambria, Dongzhan Zhou

机构 * Fudan University(复旦大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Nanyang Technological University(南洋理工大学) University of New South Wales(新南威尔士大学) National University of Singapore(新加坡国立大学) Wuhan University(武汉大学)

专题命中 领域大模型 :LLM(title_cn,summary_cn);large language model(abstract);language model(abstract);pretraining(abstract)

AI总结 本文提出ResearchBench,首个评估LLM在科学发现子任务中的基准,包含灵感检索、假设生成和排序,通过自动化框架提取跨12学科论文的关键信息,验证其准确性,并展示LLM在非分布任务中表现优异。

Comments Accepted by ACL 2026 (findings)

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.07717 2026-04-14 cs.CL cs.AI 91%

Detecting HIV-Related Stigma in Clinical Narratives Using Large Language Models

利用大语言模型检测临床叙述中的HIV相关污名化

Ziyi Chen, Yasir Khan, Mengyuan Zhang, Cheng Peng, Mengxian Lyu, Yiyang Liu, Krishna Vaddiparti, Robert L Cook, Mattia Prosperi, Yonghui Wu

机构 * Department of Health Outcomes and Biomedical Informatics, College of Medicine, University of Florida(佛罗里达大学医学院健康结果与生物医学信息学系) Department of Epidemiology, College of Public Health and Health Professions, University of Florida(佛罗里达大学公共卫生与健康专业学院流行病学系) Preston A. Wells, Jr. Center for Brain Tumor Therapy, Lillian S. Wells Department of Neurosurgery, University of Florida(佛罗里达大学莉莉安·S·威尔斯神经外科系普雷斯顿·A·威尔斯脑肿瘤治疗中心)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

AI总结 本文提出利用大语言模型开发首个能识别临床笔记中HIV污名化的NLP工具,通过专家关键词和临床词嵌入提取污名化内容,并在不同污名化子量表上评估模型性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.12476 2026-03-27 cs.CL cs.LG 91%

Retrieval-Reasoning Large Language Model-based Synthetic Clinical Trial Generation

基于检索-推理的大型语言模型合成临床试验生成

Zerui Xu, Fang Wu, Yingzhou Lu, Yuanyuan Zhang, Yue Zhao

机构 * Institute for Clarity in Documentation(清晰文档研究所) Inria Paris-Rocquencourt(巴黎- Rocquencourt 国家信息与自动化研究所) Rajiv Gandhi University(拉吉夫·甘地大学) Tsinghua University(清华大学) Palmer Research Laboratories(帕勒实验室) University of Chicago(芝加哥大学) Stanford University(斯坦福大学) Purdue University(普渡大学) University of Southern California(南加州大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

AI总结 本文提出基于检索-推理框架的合成临床试验生成方法,利用LLM生成标注二元结果的合成试验报告,通过检索模块和推理模块提升生成质量,实验证明合成数据可有效增强真实数据集并提升临床试验预测性能。

Comments Published in ACM BCB 2025. 9 pages, 4 figures, 5 tables (Main paper + Supplementary Materials)

Journal ref Proceedings of the 16th ACM International Conference on Bioinformatics, Computational Biology, and Health Informatics (ACM BCB 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.21101 2025-12-10 cs.CL cs.LG 91%

Mortgage Language Model: Domain-Adaptive Pretraining with Residual Instruction, Alignment Tuning, and Task-Specific Routing

抵押贷款语言模型:带有残差指令、对齐微调和任务特定路由的领域自适应预训练

Manish Jain, Satheesh Kumar Ponnambalam, Salman Faroz, Chandrakanth Lns, Vinay Sharma

机构 * Firstsource

专题命中 领域大模型 :language model(title,abstract);pretraining(title);LLM(abstract);large language model(abstract)

AI总结 MortgageLLM通过双专家架构和残差指令技术,在抵押贷款领域实现领域自适应预训练,提升对话问答和结构化任务性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.21748 2025-12-01 cs.CL cs.AI 91%

Building Domain-Specific Small Language Models via Guided Data Generation

通过引导数据生成构建领域专用小型语言模型

Aman Kumar, Ekant Muljibhai Amin, Xian Yeow Lee, Lasitha Vidyaratne, Ahmed K. Farahat, Dipanjan D. Ghosh, Yuta Koreeda, Chetan Gupta

专题命中 领域大模型 :language model(title,abstract);small language model(title);large language model(abstract);pretraining(abstract)

AI总结 本文提出一种通过引导数据生成构建领域专用小型语言模型的方法,结合领域适应预训练、监督微调和偏好优化,展示了在工业诊断任务中优于开源模型的性能。

Comments Accepted at Thirty-Eighth Annual Conference on Innovative Applications of Artificial Intelligence (IAAI-26)

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.14268 2025-11-25 cs.CL cs.AI cs.SI 91%

Can Large Language Models Detect Misinformation in Scientific News Reporting?

大型语言模型能否在科学新闻报道中检测虚假信息?

Yupeng Cao, Aishwarya Muralidharan Nair, Nastaran Jamalipour Soofi, Elyon Eyimife, K. P. Subbalakshmi

机构 * Stevens Institute of Technology(斯蒂文斯理工学院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

AI总结 本文探讨了使用大型语言模型检测科学新闻中虚假信息的可能性,并提出了SciNews数据集及多种基线架构进行验证。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18084 2025-11-25 cs.LG cs.AI 91%

The Alignment Paradox of Medical Large Language Models in Infertility Care: Decoupling Algorithmic Improvement from Clinical Decision-making Quality

医学大语言模型在不孕症护理中的对齐悖论:解耦算法改进与临床决策质量

Dou Liu, Ying Long, Sophia Zuoqiu, Kaipeng Xie, Runze Yang, Di Liu, Kang Li, Yiting Lin, Hanyi Liu, Rong Yin, Tian Tang

机构 * Department of Obstetrics and Gynecology, West China Second University Hospital(妇产科部门,西昌第二大学医院) Reproductive Medical Center, Department of Obstetrics and Gynecology, West China Second University Hospital(生殖医学中心,妇产科部门,西昌第二大学医院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);SFT(abstract);preference optimization(abstract)

AI总结 本研究发现,尽管GRPO在算法准确性上表现最佳,但临床医生更偏好SFT模型,因其推理更清晰且治疗可行性更高,揭示了算法改进与临床信任之间的对齐悖论。

Comments 22 pages 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.18851 2025-10-23 cs.LG cs.AI physics.chem-ph q-bio.BM q-bio.QM 91%

LICO: Large Language Models for In-Context Molecular Optimization

Tung Nguyen, Aditya Grover

机构 * Department of Computer Science University of California, Los Angeles(计算机科学系 加州大学洛杉矶分校)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);pretraining(abstract);prompting(abstract)

Comments International Conference on Learning Representations (ICLR 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17415 2025-10-21 cs.CL cs.AI cs.MA cs.MM cs.SE 91%

BenCao: An Instruction-Tuned Large Language Model for Traditional Chinese Medicine

Jiacheng Xie, Yang Yu, Yibo Chen, Hanyao Zhang, Lening Zhao, Jiaxuan He, Lei Jiang, Xiaoting Tang, Guanghui An, Dong Xu

机构 * Community Health Service Center Shanghai Pudong New Area(上海浦东新区社区卫生服务中心) School of Acupuncture-Moxibustion and Tuina, Shanghai University of Traditional Chinese Medicine(上海中医药大学针灸推拿学院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);instruction tuning(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20162 2025-09-30 cs.CL cs.AI 91%

Embedding Domain Knowledge for Large Language Models via Reinforcement Learning from Augmented Generation

Chaojun Nie, Jun Zhou, Guanxiang Wang, Shisong Wu, Zichen Wang

机构 * Laboratory of Speech and Intelligent Information Processing, Institute of Acoustics, Chinese Academy of Sciences(语音与智能信息处理实验室,声学研究所,中国科学院) University of Chinese Academy of Sciences(中国科学院大学) China Southern Power Grid Artificial Intelligence Technology Co., Ltd.(中国南方电网人工智能技术有限公司)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);post-training(abstract);SFT(abstract)

Comments Corrected author name spelling

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20381 2025-09-26 cs.CL cs.AI 91%

USB-Rec: An Effective Framework for Improving Conversational Recommendation Capability of Large Language Model

Jianyu Wen, Jingyun Wang, Cilin Yan, Jiayin Cai, Xiaolong Jiang, Ying Zhang

机构 * Harbin Institute of Technology(哈尔滨工业大学) Beihang University(北航) Xiaohongshu Inc.(小红书公司)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);preference optimization(abstract)

Comments Accepted by Recsys'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00934 2025-09-22 cs.CL cs.AI 91%

MedCOD: Enhancing English-to-Spanish Medical Translation of Large Language Models Using Enriched Chain-of-Dictionary Framework

Md Shahidul Salim, Lian Fu, Arav Adikesh Ramakrishnan, Zonghai Yao, Hong Yu

机构 * Center for Healthcare Organization and Implementation Research, VA Bedford Health Care(医疗组织与实施研究中心,VA贝德福德医疗中心) Miner School of Computer and Information Sciences, University of Massachusetts Lowell(矿物计算机与信息科学学院,马萨诸塞州立大学洛厄尔分校) Manning College of Information and Computer Sciences, University of Massachusetts Amherst(曼宁信息与计算机科学学院,马萨诸塞州立大学阿姆赫斯特分校)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

Comments To appear in Findings of the Association for Computational Linguistics: EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15192 2025-08-22 cs.AI cs.CL 91%

LLM4Sweat: A Trustworthy Large Language Model for Hyperhidrosis Support

Wenjie Lin, Jin Wei-Kocsis

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.07611 2025-07-15 cs.CL cs.AI 91%

Knowledge-Augmented Multimodal Clinical Rationale Generation for Disease Diagnosis with Small Language Models

Shuai Niu, Jing Ma, Hongzhan Lin, Liang Bai, Zhihua Wang, Yida Xu, Yunya Song, Xian Yang

机构 * Hong Kong Baptist University(香港 Baptist 大学) Shanxi University(山西大学) Shanghai Institute for Advanced Study of Zhejiang University(浙江大学上海先进研究院) Hong Kong University of Science and Technology(香港科技大学) The University of Manchester(曼彻斯特大学)

专题命中 领域大模型 :language model(title,abstract);small language model(title,abstract);LLM(abstract);large language model(abstract)

Comments 13 pages. 7 figures

Journal ref This paper is accpeted by ACL2025(Main)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.19484 2025-06-25 cs.CL cs.AI cs.HC 91%

Dialogic Pedagogy for Large Language Models: Aligning Conversational AI with Proven Theories of Learning

Russell Beale

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.12242 2025-06-17 cs.CL cs.AI cs.CY 91%

Large Language Models for History, Philosophy, and Sociology of Science: Interpretive Uses, Methodological Challenges, and Critical Perspectives

Arno Simons, Michael Zichert, Adrian Wüthrich

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);pretraining(abstract)

Comments 27 pages, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.00054 2025-05-21 cs.CL cs.AI 91%

Automating Intervention Discovery from Scientific Literature: A Progressive Ontology Prompting and Dual-LLM Framework

Yuting Hu, Dancheng Liu, Qingyun Wang, Charles Yu, Chenhui Xu, Qingxiao Zheng, Heng Ji, Jinjun Xiong

机构 * University at Buffalo(布法罗大学) University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

专题命中 领域大模型 :LLM(title,abstract);prompting(title,abstract);large language model(abstract);language model(abstract)

Comments Accepted by IJCAI2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.12738 2025-05-20 cs.LG cs.AI cs.SI 91%

EpiLLM: Unlocking the Potential of Large Language Models in Epidemic Forecasting

Chenghua Gong, Rui Sun, Yuhao Zheng, Juyuan Zhang, Tianjun Gu, Liming Pan, Linyuan Lv

机构 * University of Science and Technology of China(中国科学技术大学) East China Normal University(华东师范大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);foundation model(abstract)

Comments 18 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.01454 2025-05-13 cs.LG cs.AI stat.ME stat.ML 91%

Integrating Large Language Models in Causal Discovery: A Statistical Causal Approach

Masayuki Takayama, Tadahisa Okuda, Thong Pham, Tatsuyoshi Ikenoue, Shingo Fukuma, Shohei Shimizu, Akiyoshi Sannai

机构 * Data Science and AI Innovation Research Promotion Center, Shiga University(Shiga大学数据科学与人工智能创新研究促进中心) National Institute of Science and Technology Policy (NISTEP)(国家科学技术政策研究所) Department of Health Data Science, Tokyo Medical University Graduate School of Medicine(东京医科大学医学研究生院健康数据科学系) Human Health Sciences, Kyoto University(京都大学人类健康科学) Graduate School of Medicine Human Health Sciences, Kyoto University(京都大学医学研究生院人类健康科学) Center for Advanced Intelligence Project, RIKEN(RIKEN高级智能项目) Department of Epidemiology Infectious Disease Control and Prevention, Hiroshima University(广岛大学流行病学与传染病防控系) Graduate School of Biomedical and Health Sciences Data Science and AI Innovation Research Promotion Center, Shiga University(Shiga大学生物医学与健康科学研究生院数据科学与人工智能创新研究促进中心) SANKEN, The University of Osaka Faculty of Data Science, Shiga University(大阪大学数据科学学院,Shiga大学SANKEN) Institute for the Advanced Study of Human Biology, Kyoto University(京都大学人类生物学高级研究机构) Department of Physics, Kyoto University(京都大学物理系) Graduate School of Engineering, The University of Tokyo(东京大学工学研究生院) Research and Development Center for Large Language Models, National Institute of Informatics(信息处理研究所大型语言模型研究开发中心)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

Journal ref Published in Transactions in Machine Learning Research (05/2025) https://openreview.net/forum?id=Reh1S8rxfh

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.21591 2025-04-10 cs.AI cs.CL q-bio.GN q-bio.QM 91%

Can Large Language Models Replace Data Scientists in Biomedical Research?

Zifeng Wang, Benjamin Danek, Ziwei Yang, Zheng Chen, Jimeng Sun

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.05075 2025-01-10 cs.AI cs.LG 91%

A Text-Based Knowledge-Embedded Soft Sensing Modeling Approach for General Industrial Process Tasks Based on Large Language Model

Shuo Tong, Han Liu, Runyuan Guo, Xueqiong Tian, Wenqing Wang, Ding Liu, Youmin Zhang

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.16574 2024-10-23 cs.AI cs.LG 91%

How Can We Diagnose and Treat Bias in Large Language Models for Clinical Decision-Making?

Kenza Benkirane, Jackie Kay, Maria Perez-Ortiz

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.12858 2024-10-18 cs.CL cs.AI 91%

Large Language Models for Medical OSCE Assessment: A Novel Approach to Transcript Analysis

Ameer Hamza Shakur, Michael J. Holcomb, David Hein, Shinyoung Kang, Thomas O. Dalton, Krystle K. Campbell, Daniel J. Scott, Andrew R. Jamieson

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.11863 2024-10-17 cs.HC cs.AI cs.CL 91%

ChatVis: Automating Scientific Visualization with a Large Language Model

Tanwi Mallick, Orcun Yildiz, David Lenz, Tom Peterka

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.18995 2024-10-01 cs.CL cs.AI 91%

Systematic Characterization of the Effectiveness of Alignment in Large Language Models for Categorical Decisions

Isaac Kohane

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

Comments 19 pages (without Appendix) Appendix 7 pages. 7 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.09656 2024-08-21 cs.AI cs.CL q-bio.NC 91%

A Comparison of Large Language Model and Human Performance on Random Number Generation Tasks

Rachel M. Harrison

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12822 2024-07-19 cs.CL cs.AI 91%

Lightweight Large Language Model for Medication Enquiry: Med-Pal

Kabilan Elangovan, Jasmine Chiat Ling Ong, Liyuan Jin, Benjamin Jun Jie Seng, Yu Heng Kwan, Lit Soo Tan, Ryan Jian Zhong, Justina Koi Li Ma, YuHe Ke, Nan Liu, Kathleen M Giacomini, Daniel Shu Wei Ting

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.02833 2024-07-04 cs.IR cs.CL cs.LG 91%

LANE: Logic Alignment of Non-tuning Large Language Models and Online Recommendation Systems for Explainable Reason Generation

Hongke Zhao, Songming Zheng, Likang Wu, Bowen Yu, Jing Wang

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.18028 2024-05-29 cs.CL cs.AI 91%

Edinburgh Clinical NLP at MEDIQA-CORR 2024: Guiding Large Language Models with Hints

Aryo Pradipta Gema, Chaeeun Lee, Pasquale Minervini, Luke Daines, T. Ian Simpson, Beatrice Alex

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏